Vapi
Member of Technical Staff, Site Reliablity Engineer
San Francisco · Staff+
Sponsorship not specifiedDetected 48 days ago
TypeScriptGoBashKubernetesPrometheusGrafanaDatadogSite Reliability EngineeringCRMAuditingLoad Testing
About the role
- Most phone systems trap callers in menus and scripts.
- Vapi runs live phone calls - a p99 spike means callers drop.
- Capacity planning, load testing, and KEDA-based autoscaling for Vapi's wscaler and workerpool-cron-scaler are on your plate.
Responsibilities
- 90 Day: Ship a real platform service - capacity forecaster, auto-remediation, or oncall tooling - in Go or TypeScript. Own the postmortem process. Drive a measurable improvement in p99 call completion or MTTR.
- Build the human interface for every business
- 99.99% call completion is the number this role drives.
Requirements
- You know backpressure and autoscaling patterns - KEDA, custom metrics scaling.
Compensation
- Real stake: We offer a competitive salary and excellent equity ownership
Benefits
- Real stake: We offer a competitive salary and excellent equity ownership
- Comprehensive health coverage: medical, dental, and vision plans
- Flexible time off: take what you need
- More: catered meals, transportation, gym, and a $10k annual L&D budget
- You can build platform services in Go or TypeScript (matches Vapi's cluster-manager, database-health, wscaler, incidentManager).
- We've had 15 stability-gap outages worth learning from, and we need someone who runs incident command, owns SLOs and error budgets, and builds the reliability culture from scratch.
Company info
- Amazon Ring, ServiceTitan, New York Life, Intuit, Kavak, and thousands more, from YC startups to the Fortune 500
- 70% of the company are previous founders
This listing is sourced directly from Vapi's careers page and normalized into a canonical job model.