wagey.ggwagey.gg
30,434  jobs30,434  jobs
Browse Tech JobsCompaniesFeaturesPricingFAQs
Log InGet Started Free
Jobs(30,434)/Site Reliability Engineer Role(169)/Remote (47) - Senior Site Reliability Engineer
Remote

Remote - Senior Site Reliability Engineer

Remote - EMEA$53k - $53k1w ago
RemoteSeniorEMEACloud ComputingTransportationSite Reliability EngineerTeam LeadKubernetesDockerAWSReportingTerraformPrometheusGrafanaBashLinuxElixirPythonObservableNode.jsBack-end

Requirements

• Solid professional experience in SRE, DevOps, or Platform Engineering. • Solid hands-on Kubernetes: operating and scaling production clusters and container tooling (Docker) and its ecosystem. • Experience building and managing cloud infrastructure on AWS (or similar). • Strong infrastructure-as-code practice with Terraform. • Experience with reliability frameworks: SLOs, SLIs, error budgets, alerting strategies. • Solid observability background: OpenTelemetry, Grafana/Prometheus or similar. • Proficiency with CI/CD (GitLab CI, GitHub Actions, or similar) and deployment automation. • Comfortable with Golang, Bash/scripting; broader programming a plus. • Practical, embedded use of AI in infra/ops/dev work, agentic workflows with concrete, observable results, not just familiarity with the tools. • Clear and thoughtful communication, especially in an async-first, global setting • Proactive, curious, and comfortable taking ownership of challenges • Collaborative and respectful across cultures, time zones, and backgrounds • Experience with 1 back-end programming language (Elixir, Nodejs, Python, etc) • Experience running and configuring Linux systems in a non-cloud environment • Security knowledge and capabilities from a defensive and offensive standpoint

Responsibilities

• Lead solution discovery and delivery for reliability and infrastructure problems with real ambiguity, complexity, or scope. Autonomously, coordinating with other contributors where needed. • Contribute to the platform's architecture, tooling, and roadmap. Influence team priorities and advocate for technical initiatives. • Help define and operate reliability practices for our platform: SLOs/SLIs, error budgets, alerting, observability. Take responsibility for the team's operational stance, using support/incident metrics to shape technical strategy. • Resolve cross-team requests, identify systemic issues, and turn recurring ones into reusable fixes and runbooks rather than one-off answers. • Work AI-natively and operationalise it for the team: use agentic workflows by default; build reusable prompts, skills, and tooling embedded in the codebase so others ship faster, safely; design agent-ready systems (clean interfaces, good observability) that make AI-assisted changes easy to review. Establish shared standards and domain-level guardrails (secure-by-default patterns, CI protections, AI-assisted review practices). • Mentor and give timely, actionable feedback to less-senior engineers; participate in hiring, onboarding, and RFC discussions. • Collaborate with Security on platform hardening and threat mitigation; contribute to capacity and cost-efficiency of the infrastructure. • Participate in incident response and on-call rotations to rapidly resolve issues and maintain system reliability. • Practicals • You'll report to: SRE Team Lead • You'll report to: • Team: Engineering • Team: • Location: For this hire, due to diversity and timezones requirements, we’re prioritising Europe • Location: • Start date: As soon as possible • Start date: • Application process • Interview with recruiter • Interview with HM • (async) Infrastructure exercise (you're not expected to spend more than 2 - 4 hours) • Interview with the team (without any manager in the call so you can really get to know the people and ask all you want to ask) • Bar Raiser Interview • Offer + Background check (Veremark & Remote) • Remote's Total Rewards philosophy is to ensure fair, unbiased compensation and fair equity pay along with competitive benefits in all locations in which we operate. We do not agree to or encourage cheap-labor practices and therefore we ensure to pay above in-location rates. We hope to inspire other companies to support global talent-hiring and bring local wealth to developing countries. • At first glance our salary bands seem quite wide - here is some context. At Remote we have international operations and a globally distributed workforce.  We use geo ranges to consider geographic pay differentials as part of our global compensation strategy to remain competitive in various markets while we hiring globally.Our salary ranges are determined by role, level and location, and our job titles may span more than one career level. The actual base pay for the successful candidate in this role is dependent upon many factors such as location, transferable or job-related skills, work experience, relevant training, business needs, and market demands. The base salary range may be subject to change.At Remote, we foster internal mobility as a key element of our culture of employee growth and development, supported by a compensation philosophy that guarantees pay equity and fairness. Therefore, all compensation changes associated with an internal move will be reviewed by the Total Rewards & People Enablement team on a case by case basis. • The annual salary range for this full-time position is • $53,300—$119,850 USD

Benefits

• Our full benefits & perks are explained in our handbook at remote.com/r/benefits. As a global company, each country works differently, but some benefits/perks are for all Remoters: • work from anywhere • flexible paid time off • flexible working hours (we are async) • 16 weeks paid parental leave • mental health support services • learning budget • budget for local in-person social events or co-working spaces • How you’ll plan your day (and life) • We work async at Remote which means you can plan your schedule around your life (and not around meetings). Read more at remote.com/async. • You will be empowered to take ownership and be proactive. When in doubt you will default to action instead of waiting. Your life-work balance is important and you will be encouraged to put yourself and your family first, and fit work around your needs. • life-work balance • If that sounds like something you want, apply now!

Apply in one click

Upload My Resume

Drop here or click to browse · Tap to choose · PDF, DOCX, DOC, RTF, TXT

Apply in One Click
Apply in One Click

Similar roles

ValtechValtech - Senior Site Reliability Engineer1w ago
·Remote - Poland
RemoteEMEASeniorCloud ComputingE-commerceSite Reliability EngineerAWSGCPAzurePrometheusDynatraceDatadogNew RelicJenkinsGrafanaDockerKubernetesJavaCoaching
ValtechValtech - Senior Site Reliability Engineer1w ago
·Portugal - Remote - Hybrid
In OfficeEMEASeniorCloud ComputingE-commerceSite Reliability EngineerAWSGCPAzurePrometheusDynatraceDatadogNew RelicJenkinsGrafanaDockerKubernetesJavaCoaching
KeyrockKeyrock - SRE - Site Reliability Engineer4mo ago
·Remote - Brussels, Belgium
RemoteEMEASeniorCloud ComputingCryptocurrencySite Reliability EngineerAssociatePythonBashAWSKubernetesTerraform
replitreplit - Senior Site Reliability Engineer2mo ago
·Remote - Europe
RemoteEMEASeniorCloud ComputingSite Reliability EngineerGoPythonReportingKubernetesGCP
ClearScore Technology LimitedClearScore Technology Limited - Site Reliability Engineer1w ago
·London, England, United Kingdom
In OfficeEMEACloud ComputingSite Reliability EngineerPythonLinuxAWSKubernetesDocker SwarmNomadGoJenkinsSpinnakerElasticsearchLokiPrometheusDatadogGrafanaKafkaTerraform
patsnappatsnap - Site Reliability Engineering (SRE) Leader2d ago
·Remote - UK
RemoteEMEASeniorCloud ComputingSoftwareSite Reliability EngineerMandarinPerformance ManagementRisk ManagementAWSDockerKubernetesClaude
Miro CareersMiro Careers - Senior Network Site Reliability Engineer1mo ago
·Remote - Anywhere·Equity
RemoteEMEASeniorCloud ComputingSite Reliability EngineerSoftware EngineerTerraformAWSLinuxGovernanceBashPythonKubernetesAzureGCPReact
fundamentalfundamental - MLOps Lead4w ago
·Europe·Equity
In OfficeEMEAStaffCloud ComputingArtificial IntelligenceTeam LeadMLOpsPipeline ManagementBashPythonGoMLflowTritonKubernetesAWSGCPAzureTerraformPrometheusHelmGrafanaDatadogAirflowFastAPISnowflakeKubeflowDatabricks
talkiatrytalkiatry - Senior Site Reliability Engineer1w ago
·Remote - Americas
RemoteNASeniorCloud ComputingSite Reliability EngineerTerraformPythonTypeScriptTeam ManagementPrometheusDatadogAWSGrafanaNode.jsReactKubernetesDocumentation

Browse more by category

Show 169 moreSite Reliability EngineerShow 324 moreTeam LeadShow 1,507 moreKubernetesShow 794 moreDockerShow 2,755 moreAWSShow 6,420 moreReportingShow 820 moreTerraformShow 190 morePrometheusShow 237 moreGrafanaShow 326 moreBash
Privacy·Terms··Contact·FAQ·Wagey on X