wagey.ggwagey.gg
30,244  jobs30,244  jobs
Browse Tech JobsCompaniesFeaturesPricingFAQs
Log InGet Started Free
Jobs(30,244)/Site Reliability Engineer Role(165)/GitLab (204) - Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms
GitLab

GitLab - Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms

Remote - Canada$126k - $314k+ Equity1w ago
RemoteStaffNACloud ComputingSite Reliability EngineerTerraformKubernetesRubyGoAWSGCP

Requirements

• Experience keeping production systems reliable, combining an operations mindset with real software engineering practice • Experience building net-new infrastructure tooling and automation, not just configuring existing tools. For example, Terraform modules, Kubernetes operators or controllers, or production automation and services written from scratch • The ability to read, debug, and reason about code. Most of our teams work in Go; some work in Ruby. You can discuss a piece of code's behavior, performance, and failure modes • Experience with infrastructure as code, and with Kubernetes and its ecosystem, at a depth appropriate to your level • Hands-on experience with at least one major cloud provider (GCP or AWS) • Familiarity with observability practices, including metrics, logging, alerting, and SLOs or SLIs, and using data to inform operational decisions • Comfort participating in on-call and incident response, with a structured approach to troubleshooting under pressure • Strong written communication and the ability to operate as a manager-of-one in an async, distributed environment • A track record of using automation, and increasingly AI, to reduce toil and improve how you and your team work • Alignment with GitLab's values and a commitment to working in accordance with them • Infrastructure Platforms is responsible for the availability, reliability, performance, and scalability of GitLab's user-facing services, most notably GitLab.com. The department spans sub-departments including Production Engineering and Dedicated, and the teams within them own everything from the production fleet and networking platform to observability, incident response, and our single-tenant Dedicated offering. We are a globally distributed, all-remote group that works asynchronously, favors automation over toil, and closes the loop with monitoring and metrics to drive accountability. For more on how we work, see the Infrastructure Handbook Page. • The base salary range for this role’s listed level is currently for residents of the United States only. This range is intended to reflect the role's base salary rate in locations throughout the US. Grade level and salary ranges are determined through interviews and a review of education, experience, knowledge, skills, abilities of the applicant, equity with other team members, alignment with market data, and geographic location. The base salary range does not include any bonuses, equity, or benefits. See more information on our benefits and equity. Sales roles are also eligible for incentive pay targeted at up to 100% of the offered base salary.

Responsibilities

• Keep user-facing services and production systems reliable, scalable, and efficient • Build automation and tooling that reduces toil and replaces manual work with repeatable, infrastructure-as-code-driven workflows • Operate and troubleshoot production systems on Kubernetes, including deployments, rollouts, and scaling • Write and maintain infrastructure as code, and ship changes safely through CI/CD and GitOps • Participate in on-call, triage alerts, follow and improve runbooks, and escalate appropriately • Contribute to the observability stack, using metrics, logs, and SLOs to detect symptoms early rather than just outages • Take part in incident response and post-incident reviews, turning learnings into changes in automation and process • Document runbooks, architecture decisions, and reviews so your findings become repeatable practices

Benefits

• $126,400 - $314,400 USD • How GitLab Supports Full-Time Employees • Benefits to support your health, finances, and well-being • Flexible Paid Time Off • Team Member Resource Groups • Equity Compensation & Employee Stock Purchase Plan • Growth and Development Fund • Please note that we welcome interest from candidates with varying levels of experience; many successful candidates do not meet every single requirement. Additionally, studies have shown that people from underrepresented groups are less likely to apply to a job unless they meet every single qualification. If you're excited about this role, please apply and allow our recruiters to assess your application. • Country Hiring Guidelines: GitLab hires new team members in countries around the world. All of our roles are remote, however some roles may carry specific location-based eligibility requirements. Our Talent Acquisition team can help answer any questions about location after starting the recruiting process. • Country Hiring Guidelines:

Apply in one click

Upload My Resume

Drop here or click to browse · Tap to choose · PDF, DOCX, DOC, RTF, TXT

Apply in One Click
Apply in One Click

Similar roles

GitLabGitLab - Senior Site Reliability Engineer, Environment Automation4mo ago
·Remote - Canada·$124k - $266k/year + Equity
RemoteNASeniorCloud ComputingSite Reliability EngineerTerraformKubernetesAnsibleRubyGo
goteleportgoteleport - Senior Site Reliability Engineer - US3w ago
·Remote - United States (Remote)·$222k - $342k/year
RemoteNASeniorCloud ComputingSite Reliability EngineerGoLinuxGitHubKubernetesAWS
PinterestPinterest - Site Reliability Engineer II, tvScientific1mo ago
·San Francisco, California, United States·$114k - $114k/year + Equity
RemoteNAMidCloud ComputingSite Reliability EngineerBashPythonAWSKubernetesTerraform
Chainlink LabsChainlink Labs - Site Reliability Engineer II5mo ago
·Remote - Canada, United States, Brazil...
RemoteNAMidCryptocurrencyCloud ComputingSite Reliability EngineerGoShellPythonKubernetesTerraform
SpotifySpotify - Site Reliability Engineer5mo ago
·Remote - New York, NY·$133k - $190k/year
RemoteNACloud ComputingArtificial IntelligenceSite Reliability EngineerAWSGCPTerraformReactPython
Okta, Inc.Okta, Inc. - Staff Site Reliability Engineer1mo ago
·Bengaluru, India
In OfficeAPACStaffCloud ComputingSoftwareSite Reliability EngineerKubernetesTimeline ManagementAWSGCPHelmTerraformPythonGoGoogle GKELinuxAnsible
Clover HealthClover Health - Senior Site Reliability Engineer1w ago
·Remote - USA·$160k - $160k/year + Equity
RemoteNASeniorCloud ComputingSite Reliability EngineerGCPAzureAWS
Veeam SoftwareVeeam Software - Senior Site Reliability Engineer- FedRamp1w ago
·Remote - USA·$173k - $173k/year
RemoteNASeniorCloud ComputingArtificial IntelligenceSite Reliability EngineerGoJavaC#TypeScriptAWSAzurePrometheusGrafanaELKTerraformKubernetesDocumentationGitB2BPulumiCloseCosmosElastic Stack

Browse more by category

Show 165 moreSite Reliability EngineerShow 808 moreTerraformShow 1,490 moreKubernetesShow 346 moreRubyShow 1,653 moreGoShow 2,725 moreAWSShow 1,125 moreGCP
Privacy·Terms··Contact·FAQ·Wagey on X