Invoca
Senior ML Engineer
Remote - US & Canada (see post for locations) · Senior
No sponsorship$152k-$228kDetected 54 days ago
PythonKubernetesCI/CDMachine LearningDeep LearningPyTorchData EngineeringNLPLLMsAgentic AIMLOpsStatisticsCollaboration
About the role
- You'll be a primary driver of the infrastructure powering our Context Engine and agentic AI workflows, working closely with Data Scientists, Data Engineers, and Applied AI Engineers.
- Serve as the primary SME for operational excellence across the Invoca ML stack.
- We move quickly, swarm on hard problems, and care deeply about code quality, reliability, and each other's growth.
Responsibilities
- Architect, implement, and maintain CI/CD pipelines for ML artifacts - including model evaluation, versioning, and automated deployment.
- Own the full inference infrastructure: model serving on Triton Inference Server, Baseten, and Kubernetes-based GPU infrastructure.
- Profile and tune for low latency and high throughput, and build robust, scalable APIs for internal and external model access.
- Partner closely with Data Scientists, Data Engineers, and Applied AI Engineers to build the foundational ML systems behind Invoca's agentic AI products.
- We're hiring a Senior ML Engineer to own the productionization layer of Invoca's ML stack - model serving, inference optimization, fine-tuning, and the APIs and pipelines that tie it all together.
- Design and Optimize SLM/LLM Deployment: Own the full inference infrastructure: model serving on Triton Inference Server, Baseten, and Kubernetes-based GPU infrastructure.
- Collaborate Across Teams: Partner closely with Data Scientists, Data Engineers, and Applied AI Engineers to build the foundational ML systems behind Invoca's agentic AI products.
- Deliver Customer Value: Work with product and engineering to understand customer needs and ship ML solutions that make a measurable difference.
Requirements
- 5+ years of ML Engineering experience with a strong production focus
- Demonstrated track record deploying and maintaining transformer-based NLP models in production
- Proficiency with inference infrastructure: Triton, Baseten, vLLM, TGI, SageMaker, Vertex AI, or similar
- Familiarity with MLOps tooling, model monitoring, and eval platforms (Braintrust, MLflow, or equivalent)
- Candidates must be based within ~2 hour drive of these areas.
- Occasional business travel may be required.
- B.S. in Computer Science, Engineering, Statistics, or equivalent; advanced degree a plus
- Advanced Python and deep learning proficiency (PyTorch, HuggingFace Transformers, spaCy)
- Hands-on experience fine-tuning SLMs/LLMs (LoRA, QLoRA, PEFT) and optimizing models via quantization, batching, and throughput tuning
- Experience building production-grade APIs that expose ML models to downstream consumers
- Familiarity with RLHF or preference training is a bonus
- 📍 Location This is a remote-first role. We are currently hiring in the following locations: 📍
- United States: Greater Los Angeles Area (including Santa Barbara and San Diego) · SF Bay Area · Denver Metro · Austin Metro · Chicago Metro · Greater NYC Area
- Canada: Toronto (AI/ML technical roles only)
- Candidates must be based within ~2 hour drive of these areas. Occasional business travel may be required.
Nice to have
- B.S. in Computer Science, Engineering, Statistics, or equivalent
- advanced degree a plus
Compensation
- Salary, Benefits & Perks:
Benefits
- Flexible Time Off - We encourage a healthy work-life balance.
- Our flexible paid time off policy allows you to recharge and take time away as needed.
- Health Benefits - Our healthcare program includes medical, dental, and vision coverage, with multiple plan options so you can choose what works best for you and your family.
- Retirement - Invoca offers a 401(k) plan through Fidelity with a company match of up to 4%.
- Stock Options - All employees are invited to share in Invoca's success through stock options.
- Mental Health Program - Well-being support on a broad range of issues is available through our SpringHealth program.
- Paid Family Leave - Up to 6 weeks of 100% paid leave is provided for baby bonding, adoption, and caring for family members.
- Paid Medical Leave - Up to 12 weeks of 100% paid leave is provided for childbirth and medical needs.
- InVacation - As a thank-you to our long-term team members, we offer a bonus after 7 years of service.
- Wellness Subsidy - We provide a subsidy that can be applied toward gym memberships, fitness classes, and more.
- Salary Range $152,000 - $228,000 USD plus bonus + equity
- At Invoca, all new hires in the U.S. receive benefits starting on day one of employment.
Company info
- Senior ML Engineer
- About Invoca
- Invoca is an AI-powered revenue execution platform that brings together marketing, commerce, and contact center teams to turn every customer interaction into measurable, profitable growth.
- Join our dynamic, fast-growing team, where innovation and collaboration are at the core of our culture.
- The Data Platform team owns the full ML lifecycle at Invoca, from model training and fine-tuning through inference optimization and production APIs.
- Learn more on our blog or check out our open source projects.
Visa & Work Authorization
- Please note that we are unable to provide initial visa sponsorship for this position.
This listing is sourced directly from Invoca's careers page and normalized into a canonical job model.