wagey.ggwagey.gg
30,165  jobs30,165  jobs
Browse Tech JobsCompaniesFeaturesPricingFAQs
Log InGet Started Free
Jobs(30,165)/Site Reliability Engineer Role(165)/Kong (34) - Staff Site Reliability Engineer
Kong

Kong - Staff Site Reliability Engineer

United States$150k - $210k1mo ago
In OfficeStaffNAArtificial IntelligenceSoftwareSite Reliability EngineerPrincipalKubernetesPlaneTerraformHelmPostgreSQL

Requirements

• BS in Computer Science or equivalent; substantial experience at Staff or Principal IC level in SRE/Platform Engineering. • Proven track record building SRE or platform engineering practices for developer-facing platforms or PaaS/SaaS products — ideally at greenfield stage. • Deep Kubernetes expertise: multi-tenant cluster design, networking (CNI, service mesh, ingress), autoscaling, and security hardening. • Kong Inc., a leading developer of API and AI connectivity technologies, is building the infrastructure that powers the agentic era. Trusted by the Fortune 500 and startups alike, Kong's unified API and AI platform, Kong Konnect, enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI models. For more information, visit www.konghq.com.

Responsibilities

• Own reliability for Volcano end-to-end: Define and drive SLOs, error budgets, and incident response practices for all Volcano services — edge deployments, managed Postgres, auth, realtime, storage, and the control plane. • Own reliability for Volcano end-to-end: • Architect the platform's infrastructure: Design and build the multi-region Kubernetes infrastructure, networking, and data plane that powers Volcano's edge deployment pipeline and backend-as-a-service capabilities. • Architect the platform's infrastructure: • Build the GitOps and CI/CD backbone: Establish deployment automation, canary pipelines, and preview environment provisioning using ArgoCD, Helm, and Terraform/Terragrunt — setting patterns the broader team will follow. • Build the GitOps and CI/CD backbone: • Scale managed data services: Design, operate, and harden multi-tenant PostgreSQL clusters, Redis caching layers, and object storage — with a focus on data isolation, performance, and disaster recovery. • Scale managed data services: • Drive observability from day one: Instrument every Volcano service with meaningful SLIs; build dashboards, alerts, and runbooks using Datadog, Prometheus, and Grafana before services go live, not after incidents. • Drive observability from day one: • Lead cross-functional reliability work: Collaborate with the OCTO team, product engineering, and security to bake reliability and compliance into Volcano's architecture — not bolt it on later. • Lead cross-functional reliability work: • Set SRE culture and standards: Mentor engineers across Volcano's contributing teams on reliability principles; lead postmortems, define on-call practices, and build a blameless engineering culture. • Set SRE culture and standards: • Evaluate and adopt emerging technologies: Given Volcano's greenfield nature, evaluate and make architectural decisions on edge runtimes, serverless compute, vector databases, and AI-native infrastructure components. • Evaluate and adopt emerging technologies:

Benefits

• Kong has different base pay ranges for different work locations within the United States and Canada, which allows us to pay employees competitively and consistently in different geographic markets. Compensation varies depending on a wide array of factors, including but not limited to specific candidate location, role, skill set and level of experience. Certain roles are eligible for additional rewards including sales incentives depending on the terms of the applicable plan and role. Benefits may vary depending on location. US based employees are typically offered access to healthcare benefits, a 401(k) plan, short and long term disability benefits, basic life and AD&D insurance, among others. • Are you ready to unlock intelligence? • If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others. • Kong is building Project Volcano, an internal developer platform purpose-built for Kong's engineering ecosystem. Volcano will provide teams with on-demand preview environments, edge deployments, managed PostgreSQL, auth, realtime, and storage APIs all deeply integrated with Kong products. • As the Staff SRE for Volcano, you will be the founding reliability voice for this platform. This role is a strategic initiative driven by the Office of the CTO (OCTO). You will partner directly with engineering leadership to define the platform's reliability posture, build its SRE practice from the ground up, and ensure Volcano can scale to serve all of Kong's customers. This is a high-visibility, high-impact role with direct influence on Kong's next generation developer platform.

Apply in one click

Upload My Resume

Drop here or click to browse · Tap to choose · PDF, DOCX, DOC, RTF, TXT

Apply in One Click
Apply in One Click

Similar roles

Cambridge Mobile TelematicsCambridge Mobile Telematics - Principal Site Reliability Engineer, Machine Learning2d ago
·Cambridge, MA, US·$142k - $178k/year + Equity
In OfficeNAPrincipalCloud ComputingArtificial IntelligenceSite Reliability EngineerPrincipalPythonReportingAWSDatadogTerraformLinuxUbuntuDockerKubernetesRayDatabricks
simspace-corporationsimspace-corporation - Staff Site Reliability Engineer2mo ago
·Remote - USA·$165k - $230k/year + Equity
RemoteNAStaffSite Reliability EngineerPrincipalGoPythonKubernetesMoveKustomize
Domino Data LabDomino Data Lab - Staff Site Reliability Engineer1mo ago
·Remote - US·$200k - $200k/year
RemoteNAStaffArtificial IntelligenceSoftwareSite Reliability EngineerCloseLinuxKubernetesPythonGo
IroncladIronclad - Senior Staff Site Reliability Engineer2mo ago
·San Francisco, California, United States - Hybrid·$245k - $270k/year + Equity
In OfficeNAStaffSite Reliability EngineerKubernetesTerraformClaudeZedPulumi
PointClickCarePointClickCare - Senior Site Reliability Engineer, AI Infrastructure1mo ago
·Mississauga, Ontario - Hybrid·$139k - $155k/year
In OfficeNASeniorLife SciencesSoftwareCybersecuritySite Reliability EngineerDocumentationTerraformAzureKubernetesDatabricks
AccelaAccela - Principal Site Reliability Engineer1mo ago
·Remote - Based - US·$160k - $185k/year + Equity
RemoteNAPrincipalInsuranceCloud ComputingSite Reliability EngineerPrincipalBashPythonKubernetesAzureChange Management
PlaidPlaid - Staff Site Reliability Engineer, Release Engineering1mo ago
·New York, United States·$208k - $274k/year + Equity
In OfficeNAStaffBankingSoftwareSite Reliability EngineerPlaidGoKubernetesPrometheus
GitLabGitLab - Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms1w ago
·Remote - Canada·$126k - $314k/year + Equity
RemoteNAStaffCloud ComputingSite Reliability EngineerTerraformKubernetesRubyGoAWSGCP

Browse more by category

Show 165 moreSite Reliability EngineerShow 761 morePrincipalShow 1,490 moreKubernetesShow 72 morePlaneShow 808 moreTerraformShow 111 moreHelmShow 564 morePostgreSQL
Privacy·Terms··Contact·FAQ·Wagey on X