Databricks

Databricks

Sr. Engineering Manager, AI Runtime

Mountain View, California; San Francisco, California · Senior

Sponsorship not specified$229k-$297kDetected 4 days ago
Node.jsDatabricksDeep LearningPyTorchSparkNLPLLMsElectrical EngineeringResearchLeadershipCommunicationCollaboration

About the role

  • Databricks' AI Runtime (AIR) product provides enterprises with an API for training and fine-tuning deep learning and LLM models with on-demand GPUs.
  • The pay range(s) for this role is listed below and represents the expected salary range for non-commissionable roles or on-target earnings for commissionable roles.
  • Based on the factors above, Databricks anticipates utilizing the full width of the range.

Responsibilities

  • As a Senior Engineering Manager, you will lead the team owning both the product experience and the foundational infrastructure of AIR.
  • Lead, mentor, and grow a high-performing engineering team responsible for the Custom Training product and its foundational infrastructure, including distributed training orchestration, cluster lifecycle, fault tolerance, and training efficiency.
  • Define and own the product and technical roadmap for AIR, balancing customer experience, functionality, and foundational investments.
  • Drive architectural decisions and product design for managed GPU training at scale.
  • Build observability and reliability practices for long-running, multi-node training jobs, including checkpoint strategies, failure recovery, and operational runbooks.
  • Partner with recruiting to attract, hire, and develop top-tier engineering talent.
  • Experience building platform products with clear SLAs where you've owned the customer experience, not just the backend.
  • Strong cross-functional leadership across platform, product, and research teams, with the ability to lead through ambiguity and deliver complex projects.

Requirements

  • 8+ years of software engineering experience, with 3+ years in engineering management.

Compensation

  • Databricks is committed to fair and equitable compensation practices.
  • The pay range(s) for this role is listed below and represents the expected salary range for non-commissionable roles or on-target earnings for commissionable roles.
  • Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to job-related skills, depth of experience, relevant certifications and training, and specific work location.
  • The total compensation package for this position may also include eligibility for annual performance bonus, equity, and the benefits listed above.
  • Local Pay Range
  • $228,600 - $297,120 USD

Benefits

  • At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees.
  • For specific details on the benefits offered in your region click here.
  • At Databricks, we are passionate about enabling data teams to solve the world's toughest problems, from making the next mode of transportation a reality to accelerating the development of medical breakthroughs.

Company info

  • We do this by building and running the world's best data and AI infrastructure platform so our customers can use deep data insights to improve their business.
  • Whether it's a transformer model for drug discovery or a fine-tuned foundation model, customers use this team's training infrastructure to build state-of-the-art frontier models.
  • Collaborate closely with product, research, platform, infrastructure teams, and customers to drive end-to-end delivery, from ideation and prioritization to launch and operation.
  • At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel.

Equal opportunity

  • Our Commitment to Diversity and Inclusion

This listing is sourced directly from Databricks's careers page and normalized into a canonical job model.