Harvey
Senior Engineering Manager, Production Engineering
San Francisco · Senior
Sponsorship not specified$272k-$355kDetected 8 days ago
Distributed SystemsAWSGCPAzureCloud PlatformsKubernetesTerraformTemporalMachine LearningLLMsAgentic AIAI OrchestrationCybersecurityNetwork SecurityIncident ResponseForecastingAuditingLeadershipCommunicationDemand Planning
About the role
- Our infrastructure is the foundation that powers every customer interaction, every model inference, and every production workload.
- You'll report to the Head of Infrastructure and play a key leadership role in shaping the future of Harvey's infrastructure platform.
- At Harvey, we value Decisiveness, Simplicity, and the belief that Job's Not Finished.
Responsibilities
- Lead, mentor, and grow a team of high-performing infrastructure engineers responsible for Harvey's production infrastructure foundation.
- Partner with Engineering, Security, Product, and AI Infrastructure leaders to define long-term infrastructure strategy and execution priorities.
- Drive technical direction for compute infrastructure, networking, Kubernetes, workflow orchestration, and production operations.
- Lead cross-functional initiatives to improve reliability, scalability, security, operational efficiency, and infrastructure cost optimization.
- Own and operate Harvey's global compute and network infrastructure, ensuring high availability, scalability, reliability, and performance.
- Manage compute resources to maximize utilization, performance, and service availability while supporting rapidly growing AI workloads.
- Lead capacity planning, demand forecasting, and fleet lifecycle management to ensure infrastructure scales efficiently with business growth.
- Own Harvey's Temporal-based workflow orchestration platform, ensuring reliable, scalable, and observable execution of distributed application workflows.
- Drive infrastructure cost optimization through capacity management, resource rightsizing, workload efficiency improvements, and utilization monitoring.
- Build and maintain secure infrastructure foundations, including identity and access management, network isolation, secrets management, auditing, and compliance controls.
Requirements
- Experience with infrastructure automation and Infrastructure-as-Code using tools such as Terraform or Pulumi.
- Strong understanding of compute infrastructure, networking, capacity planning, fleet management, and production operations.
- Strong understanding of infrastructure security, including IAM, network security, secrets management, and compliance best practices.
Nice to have
- Experience operating workflow orchestration platforms such as Temporal.
- Experience supporting AI/ML or LLM infrastructure at scale.
- Experience managing GPU fleets, high-performance compute infrastructure, or large-scale capacity planning.
- $272,000 - $355,000
- YOU CAN FIND ALL OF OUR APPLICANT PRIVACY NOTICES [HERE https://www.notion.so/harveyai/Harvey-Candidate-Privacy-Policies-319ac3fcdd7a803bb807d5094f249922].
Compensation
- $272,000 - $355,000
This listing is sourced directly from Harvey's careers page and normalized into a canonical job model.