Staff Site Reliability Engineer - Volcano
Kong — Tracked from its ashby job board
About the role
Are you ready to unlock intelligence?
If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.
About The Role:
Kong is building Project Volcano, an internal developer platform purpose-built for Kong's engineering ecosystem. Volcano will provide teams with on-demand preview environments, edge deployments, managed PostgreSQL, auth, realtime, and storage APIs all deeply integrated with Kong products.
As the Staff SRE for Volcano, you will be the founding reliability voice for this platform. This role is a strategic initiative driven by the Office of the CTO (OCTO). You will partner directly with engineering leadership to define the platform's reliability posture, build its SRE practice from the ground up, and ensure Volcano can scale to serve all of Kong's customers. This is a high-visibility, high-impact role with direct influence on Kong's next generation developer platform.
What You'll Do:
- Own reliability for Volcano end-to-end: Define and drive SLOs, error budgets, and incident response practices for all Volcano services — edge deployments, managed Postgres, auth, realtime, storage, and the control plane.
- Architect the platform's infrastructure: Design and build the multi-region Kubernetes infrastructure, networking, and data plane that powers Volcano's edge deployment pipeline and backend-as-a-service capabilities.
- Build the GitOps and CI/CD backbone: Establish deployment automation, canary pipelines, and preview environment provisioning using ArgoCD, Helm, and Terraform/Terragrunt — setting patterns the broader team will follow.
- Scale managed data services: Design, operate, and harden multi-tenant PostgreSQL clusters, Redis caching layers, and object storage — with a focus on data isolation, performance, and disaster recovery.
- Drive observability from day one: Instrument every Volcano service with meaningful SLIs; build dashboards, alerts, and runbooks using Datadog, Prometheus, and Grafana before services go live, not after incidents.
- Lead cross-functional reliability work: Collaborate with the OCTO team, product engineering, and security to bake reliability and compliance into Volcano's architecture — not bolt it on later.
- Set SRE culture and standards: Mentor engineers across Volcano's contributing teams on reliability principles; lead postmortems, define on-call practices, and build a blameless engineering culture.
- Evaluate and adopt emerging technologies: Given Volcano's greenfield nature, evaluate and make architectural decisions on edge runtimes, serverless compute, vector databases, and AI-native infrastructure components.
What You'll Bring:
- BS in Computer Science or equivalent; substantial experience at Staff or Principal IC level in SRE/Platform Engineering.
- Proven track record building SRE or platform engineering practices for developer-facing platforms or PaaS/SaaS products — ideally at greenfield stage.
- Deep Kubernetes expertise: multi-tenant cluster design, networking (CNI, service mesh, ingress), autoscaling, and security hardening.
#LI-BR2
About Kong:
Kong Inc., a leading developer of API and AI connectivity technologies, is building the infrastructure that powers the agentic era. Trusted by the Fortune 500 and startups alike, Kong's unified API and AI platform, Kong Konnect, enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI models. For more information, visit www.konghq.com http://www.konghq.com.
What the index says about this role
- First seen by JobLarper — Aug 2, 2026, 6 days ago. Older postings collect hundreds of applicants — a tailored résumé matters more the longer a role has been live.
- What DevOps & Site Reliability Engineer roles ask for — across 1,405 indexed openings: Cloud (50%), Python (41%), REST/APIs (33%), Go (31%), Backend (24%). This posting names Cloud, REST/APIs, Backend, AI/LLM.
- Kong is hiring actively — 45 open roles indexed.
Derived from the 26,996 roles JobLarper indexes daily from official company boards — not from the job description above.
More open roles at Kong
- Staff Curriculum DeveloperBangalore, India · Senior+
- Site Reliability EngineerMilan, Italy · Mid
- Staff Solutions Engineer (UK)London, England, United Kingdom · Senior+
- Staff Solutions Engineer - New YorkNew York, United States · Senior+ · $200K – $260K • Offers Equity • Offers Commission
- Senior Software Engineering Manager, Managed Gateways SREsWashington, United States · Manager · $153K – $218K
- Senior SRE, Managed GatewaysWashington, United States · Senior+ · $113K – $162K
- Senior Software Engineer - AI GatewayShanghai, China · Senior+
- Software Engineer, Core PlatformCanada · Mid · CA$105K – CA$195K • Offers Equity • Offers Bonus
Similar DevOps & Site Reliability Engineer roles at other companies
- Finance Systems Engineer, TaxAnthropic · San Francisco, CA | Seattle, WA
- Senior Site Reliability Engineer (SRE & Platform Reliability)Affirm · Remote Poland · Remote
- Lead Cloud DevOps Engineer (Oakland, CA Office)Fictiv · Oakland, CA Office
- Senior Application Support Engineer / Site Reliability Engineer (SRE)DTCC · Boston, MA, United States
- Senior Application Support Engineer (SRE)DTCC · Tampa, FL, United States
- Epic Site Reliability Engineer IIQuest Diagnostics · Secaucus, NJ, United States
- Android Systems Engineer, Consumer DevicesOpenAI · San Francisco
- Model Based Systems Engineer (Top Secret Clearance or Higher)IERUS Technologies · Huntsville, AL
Browse all devops & site reliability engineer jobs in united states — 721 open roles across 306 companies.