Fortive
Sr. Director, Platform & AI Infrastructure
Remote, United States · Director
Sponsorship not specified$134k-$249kDetected 13 days ago
SQLMongoDBVector DatabasesAzureCloud PlatformsSite Reliability EngineeringPlatform EngineeringMachine LearningLLMsRAGMLOpsIncident ResponseProcurementZero TrustLeadershipCommunication
About the role
- Infrastructure-as-code, capacity planning, and reliability across CPU and GPU workloads.
- Operations, performance, HA/DR, and roadmap for Oracle, SQL Server, MongoDB, and similar.
- Brief executives on reliability and risk.
Responsibilities
- Own the AI/ML platform: GPU capacity strategy, model serving and inference latency, training and fine-tuning infrastructure, MLOps and evaluation pipelines, vector and feature stores, and the RAG and agentic patterns our product teams build on.
- Partner with product engineering and architecture on build-vs-buy decisions across foundation model providers and open-source.
- Build out the incident response program: on-call structure, severity definitions, incident command, communication standards, postmortems, and follow-through on systemic fixes.
- Develop the observability stack across metrics, logs, traces, and synthetics.
- Lead internal communication during incidents.
- Lead a globally distributed team of managers and senior ICs.
- Maintain a strong culture and leadership bench.
- Familiarity with the modern AI stack: vector databases, RAG, agent frameworks, evaluation, and the build-vs-buy tradeoffs across foundation model providers and open-source.
- About Gordian Gordian is the world's leading provider of facility and construction cost data, software and services for all phases of the building lifecycle.
- From planning to design, procurement, construction and operations, Gordian's solutions help clients maximize efficiency, optimize cost savings and increase building quality.
Requirements
- Required 10+ years in platform engineering, SRE, infrastructure, or AI/ML infrastructure, with 5+ leading teams.
- Production experience running ML/AI workloads at scale, including GPU infrastructure, model serving, MLOps, or LLM/inference platforms.
Skills
- cloud infrastructure, data platforms, ML/AI infrastructure, incident response, and observability.
- AI platform and production ML.
Compensation
- The salary range for this position (in local currency) is 133,600.00 - 248,500.00
- Pay Range The salary range for this position (in local currency) is 133,600.00 - 248,500.00
Benefits
- Bonus or Equity This position is also eligible for bonus as part of the total compensation package.
- We accelerate transformation in high-impact fields like workplace safety, build environments, and healthcare.
- Our forward-looking companies lead the way in healthcare sterilization, industrial safety, predictive maintenance, and other mission-critical solutions.
Company info
- Speak to customers when major incidents affect them.
- Effective communicator with executives, the Board, customers, and the company during incidents.
- We are a global industrial technology innovator with a startup spirit.
- We're a force for progress, working alongside our customers and partners to solve challenges on a global scale, from workplace safety in the most demanding conditions to advanced technologies that help providers focus on exceptional patient care.
Equal opportunity
- We Are an Equal Opportunity Employer.
- equal opportunity employers.
- Individuals who need a reasonable accommodation because of a disability for any part of the employment
Apply directly at Fortive →Create a free account for alerts like thisView Fortive immigration profile
This listing is sourced directly from Fortive's careers page and normalized into a canonical job model.