Kong

Kong

Staff Site Reliability Engineer - Volcano

United States · Staff+

Sponsorship not specifiedDetected 30 days ago
PostgreSQLRedisVector DatabasesKubernetesTerraformHelmCI/CDPrometheusGrafanaDatadogSite Reliability EngineeringPlatform EngineeringIncident ResponseLeadership

About the role

  • Volcano will provide teams with on-demand preview environments, edge deployments, managed PostgreSQL, auth, realtime, and storage APIs all deeply integrated with Kong products.
  • As the Staff SRE for Volcano, you will be the founding reliability voice for this platform.
  • This role is a strategic initiative driven by the Office of the CTO (OCTO).

Responsibilities

  • Own reliability for Volcano end-to-end: Define and drive SLOs, error budgets, and incident response practices for all Volcano services - edge deployments, managed Postgres, auth, realtime, storage, and the control plane.
  • Architect the platform's infrastructure: Design and build the multi-region Kubernetes infrastructure, networking, and data plane that powers Volcano's edge deployment pipeline and backend-as-a-service capabilities.
  • Build the GitOps and CI/CD backbone: Establish deployment automation, canary pipelines, and preview environment provisioning using ArgoCD, Helm, and Terraform/Terragrunt - setting patterns the broader team will follow.
  • Scale managed data services: Design, operate, and harden multi-tenant PostgreSQL clusters, Redis caching layers, and object storage - with a focus on data isolation, performance, and disaster recovery.
  • Drive observability from day one: Instrument every Volcano service with meaningful SLIs
  • build dashboards, alerts, and runbooks using Datadog, Prometheus, and Grafana before services go live, not after incidents.
  • Lead cross-functional reliability work: Collaborate with the OCTO team, product engineering, and security to bake reliability and compliance into Volcano's architecture - not bolt it on later.
  • lead postmortems, define on-call practices, and build a blameless engineering culture.
  • Proven track record building SRE or platform engineering practices for developer-facing platforms or PaaS/SaaS products - ideally at greenfield stage.

Nice to have

  • BS in Computer Science or equivalent
  • substantial experience at Staff or Principal IC level in SRE/Platform Engineering.

Skills

  • What You'll Bring:
  • BS in Computer Science or equivalent; substantial experience at Staff or Principal IC level in SRE/Platform Engineering.
  • For more information, visit www.konghq.com http://www.konghq.com.

Company info

  • Mentor engineers across Volcano's contributing teams on reliability principles; lead postmortems, define on-call practices, and build a blameless engineering culture.
  • Nobody checks every box - we're looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.

This listing is sourced directly from Kong's careers page and normalized into a canonical job model.