Vapi

Vapi

Member of Technical Staff, Site Reliablity Engineer

San Francisco · Staff+

Sponsorship not specifiedDetected 48 days ago
TypeScriptGoBashKubernetesPrometheusGrafanaDatadogSite Reliability EngineeringCRMAuditingLoad Testing

About the role

  • Most phone systems trap callers in menus and scripts.
  • Vapi runs live phone calls - a p99 spike means callers drop.
  • Capacity planning, load testing, and KEDA-based autoscaling for Vapi's wscaler and workerpool-cron-scaler are on your plate.

Responsibilities

  • 90 Day: Ship a real platform service - capacity forecaster, auto-remediation, or oncall tooling - in Go or TypeScript. Own the postmortem process. Drive a measurable improvement in p99 call completion or MTTR.
  • Build the human interface for every business
  • 99.99% call completion is the number this role drives.

Requirements

  • You know backpressure and autoscaling patterns - KEDA, custom metrics scaling.

Compensation

  • Real stake: We offer a competitive salary and excellent equity ownership

Benefits

  • Real stake: We offer a competitive salary and excellent equity ownership
  • Comprehensive health coverage: medical, dental, and vision plans
  • Flexible time off: take what you need
  • More: catered meals, transportation, gym, and a $10k annual L&D budget
  • You can build platform services in Go or TypeScript (matches Vapi's cluster-manager, database-health, wscaler, incidentManager).
  • We've had 15 stability-gap outages worth learning from, and we need someone who runs incident command, owns SLOs and error budgets, and builds the reliability culture from scratch.

Company info

  • Amazon Ring, ServiceTitan, New York Life, Intuit, Kavak, and thousands more, from YC startups to the Fortune 500
  • 70% of the company are previous founders

This listing is sourced directly from Vapi's careers page and normalized into a canonical job model.