Benchling
Engineering Leader, Infrastructure
San Francisco, CA
Sponsorship not specifiedDetected 1 day ago
TypeScriptPythonAWSCloud PlatformsKubernetesDatadogSite Reliability EngineeringPlatform EngineeringAgentic AICybersecurityIncident ResponsePerformance ManagementLeadershipCollaborationMentoringGxP
About the role
- You'll work closely across Infrastructure and Platform.
- You'll be partnering with product engineering teams and other leaders to improve reliability, scalability, security, and developer productivity in a regulated environment.
- We're specifically looking for a leader with a software engineering and infrastructure background, not just operations.
Responsibilities
- Lead and grow a team of infrastructure engineers including hiring, coaching, performance management, and career development across all levels of the team.
- Own delivery and outcomes for critical infrastructure and services-platform initiatives including reliability, scalability, cost, security posture, and compliance-aligned engineering practices.
- Drive reliability improvements using pragmatic SRE principles: service-level thinking (SLIs/SLOs), operational readiness, automation, and resiliency patterns.
- Build and evolve cloud foundations on AWS, including network and compute patterns and Kubernetes-based services.
- Partner cross-functionally with Platform Engineering, Security, and Product Engineering to align priorities and deliver shared roadmaps.
- Production experience operating on AWS (e.g., EKS/ECS, RDS, VPC, EC2, S3) and building scalable internal platforms.
- Experience building or evolving observability and operational tooling (Datadog, FireHydrant, Sentry, or similar).
- Ability to collaborate across teams and influence without authority
- Ability to collaborate across teams and influence without authority; strong written and verbal communication.
Requirements
- 3+ years of people leadership experience including managing a team across a range of levels.
- Strong technical depth with previous direct, hands-on experience in software engineering and infrastructure/platform engineering.
- Strong experience with Kubernetes and services-platform patterns (e.g., service mesh such as Istio, ingress, service discovery, workload isolation).
- Comfort working in a stack largely in Python (Go also highly appreciated), with TypeScript in the broader ecosystem (especially for platform tooling and integrations).
- Direct ownership of incident management programs (on-call design, incident command, postmortems, and reliability governance).
- Experience in regulated environments (GxP, biotech, healthcare, finance), including change control, auditability, and SDLC/process rigor.
Nice to have
- Experience managing or partnering closely with SRE teams.
- Benchling welcomes everyone.
- We believe diversity enriches our team so we hire people with a wide range of identities, backgrounds, and experiences.
- We also consider for employment qualified applicants with arrest and conviction records, consistent with applicable federal, state and local law, including but not limited to the San Francisco Fair Chance Ordinance.
Skills
- When a breakthrough is delayed, the world waits.
- Getting a molecule from discovery to patients, or a crop from lab to field, involves thousands of slow, manual, disconnected steps.
- AI has the potential to change this, compressing decades of R&D work into years.
- But that only happens when clean, structured scientific data and AI are built into how science gets done.
- Benchling is the AI platform for biotech R&D.
Benefits
- Strengthen incident response: improve on-call health, incident processes, postmortems, and follow-up execution in collaboration with product engineering and platform teams.
Apply directly at Benchling →Create a free account for alerts like thisView Benchling immigration profile
This listing is sourced directly from Benchling's careers page and normalized into a canonical job model.