Braintrust
Cloud Infrastructure Engineer
San Francisco
Sponsorship not specifiedDetected 468 days ago
TypeScriptPythonAWSGCPAzureCloud PlatformsKubernetesTerraformCI/CDDevOpsSite Reliability Engineering
About the role
- This is a high-impact role where you'll contribute across our internal AWS environment and help customers deploy our stack in AWS, Azure, and GCP.
Responsibilities
- Build and maintain Terraform modules for both internal infrastructure and customer deployments
- Own and improve our CI/CD pipeline: reduce build times, improve failure visibility, and enable safer, faster releases
- Partner with engineering teams to build and evolve a secure, developer-friendly infrastructure platform
- Implement tools and automation to improve deployment, rollback, and infrastructure reliability
Nice to have
- experience with multi-cloud environments or self-hosted enterprise software
- Centralize and scale observability - including logs, metrics, dashboards, and alerts
- 5+ years of experience in DevOps, SRE, or Infrastructure Engineering roles
- Deep experience with Terraform and at least one major cloud provider (AWS strongly preferred)
Skills
- Proficient in scripting or programming (Python, Typescript, or Go)
- Experience supporting production systems and responding to incidents
- deploying, debugging, and scaling real workloads
- Bonus: experience with multi-cloud environments or self-hosted enterprise software
- Medical, dental, and vision insurance
- Daily lunch, snacks, and beverages
- Flexible time off
- Competitive salary and equity
- Wifi & cellphone stipend
- Braintrust is an equal opportunity employer.
Compensation
- Competitive salary and equity
Company info
- Braintrust is the agent observability platform.
- By actively applying intelligence to agent traces and automatically surfacing the most critical patterns, Braintrust gives teams the visibility to understand how agents behave in production and the tools to improve them.
- Teams at Notion, Stripe, Box, OpenAI, and Cloudflare use Braintrust to trace their agents, find the issues in their observability data, and run evals that tell them how to improve.
- Work directly with customers in Slack to support self-hosting and troubleshoot infrastructure issues. Build tools to make it easier for them to support themselves.
- Support multi-cloud deployment patterns (AWS primarily, with Azure and GCP support for enterprise customers)
- Comfortable working directly with customers in a support or deployment context
Equal opportunity
- All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status.
Apply directly at Braintrust →Create a free account for alerts like thisView Braintrust immigration profile
This listing is sourced directly from Braintrust's careers page and normalized into a canonical job model.