Kong
Staff Site Reliability Engineer - Volcano
United States · Staff+
Sponsorship not specifiedDetected 30 days ago
PostgreSQLRedisVector DatabasesKubernetesTerraformHelmCI/CDPrometheusGrafanaDatadogSite Reliability EngineeringPlatform EngineeringIncident ResponseLeadership
About the role
- Volcano will provide teams with on-demand preview environments, edge deployments, managed PostgreSQL, auth, realtime, and storage APIs all deeply integrated with Kong products.
- As the Staff SRE for Volcano, you will be the founding reliability voice for this platform.
- This role is a strategic initiative driven by the Office of the CTO (OCTO).
Responsibilities
- Own reliability for Volcano end-to-end: Define and drive SLOs, error budgets, and incident response practices for all Volcano services - edge deployments, managed Postgres, auth, realtime, storage, and the control plane.
- Architect the platform's infrastructure: Design and build the multi-region Kubernetes infrastructure, networking, and data plane that powers Volcano's edge deployment pipeline and backend-as-a-service capabilities.
- Build the GitOps and CI/CD backbone: Establish deployment automation, canary pipelines, and preview environment provisioning using ArgoCD, Helm, and Terraform/Terragrunt - setting patterns the broader team will follow.
- Scale managed data services: Design, operate, and harden multi-tenant PostgreSQL clusters, Redis caching layers, and object storage - with a focus on data isolation, performance, and disaster recovery.
- Drive observability from day one: Instrument every Volcano service with meaningful SLIs
- build dashboards, alerts, and runbooks using Datadog, Prometheus, and Grafana before services go live, not after incidents.
- Lead cross-functional reliability work: Collaborate with the OCTO team, product engineering, and security to bake reliability and compliance into Volcano's architecture - not bolt it on later.
- lead postmortems, define on-call practices, and build a blameless engineering culture.
- Proven track record building SRE or platform engineering practices for developer-facing platforms or PaaS/SaaS products - ideally at greenfield stage.
Nice to have
- BS in Computer Science or equivalent
- substantial experience at Staff or Principal IC level in SRE/Platform Engineering.
Skills
- What You'll Bring:
- BS in Computer Science or equivalent; substantial experience at Staff or Principal IC level in SRE/Platform Engineering.
- For more information, visit www.konghq.com http://www.konghq.com.
Company info
- Mentor engineers across Volcano's contributing teams on reliability principles; lead postmortems, define on-call practices, and build a blameless engineering culture.
- Nobody checks every box - we're looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.
This listing is sourced directly from Kong's careers page and normalized into a canonical job model.