Senior Site Reliability Engineer - Volcano
United States · Remote · Full-time
- Posted 2w ago
- From Kong’s careers page
- Location
- United States
- Work mode
- Remote
- Type
- Full-time
- Level
- Senior
- Experience
- 4+ years
- Department
- Information Technology
Opens the listing on jobs.ashbyhq.com
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.
About the role
About The Role:
Kong is building Project Volcano, an internal developer platform purpose-built for Kong's engineering ecosystem. Volcano will provide teams with on-demand preview environments, edge deployments, managed PostgreSQL, auth, realtime, and storage APIs all deeply integrated with Kong products.
As the Senior SRE for Volcano, you will be the reliability voice for this platform. This role is a strategic initiative driven by the Office of the CTO (OCTO). You will partner directly with engineering leadership to define the platform's reliability posture. This is a high-visibility, high-impact role with direct influence on Kong's next generation developer platform.
What You'll Do:
-
Own reliability for Volcano end-to-end: Define and drive SLOs, error budgets, and incident response practices for all Volcano services — edge deployments, managed Postgres, auth, realtime, storage, and the control plane.
-
Contribute to the platform's infrastructure: Design and build the multi-region Kubernetes infrastructure, networking, and data plane that powers Volcano's edge deployment pipeline and backend-as-a-service capabilities.
-
Build the GitOps and CI/CD backbone: Establish deployment automation, canary pipelines, and preview environment provisioning using ArgoCD, Helm, and Terraform/Terragrunt — setting patterns the broader team will follow.
-
Scale managed data services: Design, operate, and harden multi-tenant PostgreSQL clusters, Redis caching layers, and object storage — with a focus on data isolation, performance, and disaster recovery.
-
Drive observability from day one: Instrument every Volcano service with meaningful SLIs; build dashboards, alerts, and runbooks using Datadog, Prometheus, and Grafana before services go live, not after incidents.
-
Lead cross-functional reliability work: Collaborate with the OCTO team, product engineering, and security to bake reliability and compliance into Volcano's architecture — not bolt it on later.
-
Evaluate and adopt emerging technologies: Given Volcano's greenfield nature, evaluate and make architectural decisions on edge runtimes, serverless compute, vector databases, and AI-native infrastructure components.
What You'll Bring:
-
BS in Computer Science or equivalent; substantial experience at Staff or Principal IC level in SRE/Platform Engineering.
-
Proven track record building SRE or platform engineering practices for developer-facing platforms or PaaS/SaaS products — ideally at greenfield stage.
-
Kubernetes expertise: multi-tenant cluster design, networking (CNI, service mesh, ingress), autoscaling, and security hardening.
Skills they ask for
Pick one to see other roles that ask for it.
About Kong
API and AI connectivity platformKong provides API and AI connectivity software for managing APIs, services, events and AI applications.
See all 30 roles at KongMore roles at Kong
See all 30- Staff Product Manager, AI ManagementSan Francisco · Staff · On-siteProduct Management · Staff · On-siteSan Francisco, United States5h
- Product Manager, Billing (Senior Staff)Albany · Staff · RemoteProduct Management · Staff · RemoteAlbany, United States13h
- Senior Technical Customer Success ManagerSacramento · Senior · RemoteCustomer Service · Senior · RemoteSacramento, United States1d
- Enterprise Account Executive - NorthwestOlympia · RemoteSales · RemoteOlympia, United States3d
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.