Invoca

Invoca

Senior ML Engineer

Remote - US & Canada (see post for locations) · Senior

No sponsorship$152k-$228kDetected 54 days ago
PythonKubernetesCI/CDMachine LearningDeep LearningPyTorchData EngineeringNLPLLMsAgentic AIMLOpsStatisticsCollaboration

About the role

  • You'll be a primary driver of the infrastructure powering our Context Engine and agentic AI workflows, working closely with Data Scientists, Data Engineers, and Applied AI Engineers.
  • Serve as the primary SME for operational excellence across the Invoca ML stack.
  • We move quickly, swarm on hard problems, and care deeply about code quality, reliability, and each other's growth.

Responsibilities

  • Architect, implement, and maintain CI/CD pipelines for ML artifacts - including model evaluation, versioning, and automated deployment.
  • Own the full inference infrastructure: model serving on Triton Inference Server, Baseten, and Kubernetes-based GPU infrastructure.
  • Profile and tune for low latency and high throughput, and build robust, scalable APIs for internal and external model access.
  • Partner closely with Data Scientists, Data Engineers, and Applied AI Engineers to build the foundational ML systems behind Invoca's agentic AI products.
  • We're hiring a Senior ML Engineer to own the productionization layer of Invoca's ML stack - model serving, inference optimization, fine-tuning, and the APIs and pipelines that tie it all together.
  • Design and Optimize SLM/LLM Deployment: Own the full inference infrastructure: model serving on Triton Inference Server, Baseten, and Kubernetes-based GPU infrastructure.
  • Collaborate Across Teams: Partner closely with Data Scientists, Data Engineers, and Applied AI Engineers to build the foundational ML systems behind Invoca's agentic AI products.
  • Deliver Customer Value: Work with product and engineering to understand customer needs and ship ML solutions that make a measurable difference.

Requirements

  • 5+ years of ML Engineering experience with a strong production focus
  • Demonstrated track record deploying and maintaining transformer-based NLP models in production
  • Proficiency with inference infrastructure: Triton, Baseten, vLLM, TGI, SageMaker, Vertex AI, or similar
  • Familiarity with MLOps tooling, model monitoring, and eval platforms (Braintrust, MLflow, or equivalent)
  • Candidates must be based within ~2 hour drive of these areas.
  • Occasional business travel may be required.
  • B.S. in Computer Science, Engineering, Statistics, or equivalent; advanced degree a plus
  • Advanced Python and deep learning proficiency (PyTorch, HuggingFace Transformers, spaCy)
  • Hands-on experience fine-tuning SLMs/LLMs (LoRA, QLoRA, PEFT) and optimizing models via quantization, batching, and throughput tuning
  • Experience building production-grade APIs that expose ML models to downstream consumers
  • Familiarity with RLHF or preference training is a bonus
  • 📍 Location This is a remote-first role. We are currently hiring in the following locations: 📍
  • United States: Greater Los Angeles Area (including Santa Barbara and San Diego) · SF Bay Area · Denver Metro · Austin Metro · Chicago Metro · Greater NYC Area
  • Canada: Toronto (AI/ML technical roles only)
  • Candidates must be based within ~2 hour drive of these areas. Occasional business travel may be required.

Nice to have

  • B.S. in Computer Science, Engineering, Statistics, or equivalent
  • advanced degree a plus

Compensation

  • Salary, Benefits & Perks:

Benefits

  • Flexible Time Off - We encourage a healthy work-life balance.
  • Our flexible paid time off policy allows you to recharge and take time away as needed.
  • Health Benefits - Our healthcare program includes medical, dental, and vision coverage, with multiple plan options so you can choose what works best for you and your family.
  • Retirement - Invoca offers a 401(k) plan through Fidelity with a company match of up to 4%.
  • Stock Options - All employees are invited to share in Invoca's success through stock options.
  • Mental Health Program - Well-being support on a broad range of issues is available through our SpringHealth program.
  • Paid Family Leave - Up to 6 weeks of 100% paid leave is provided for baby bonding, adoption, and caring for family members.
  • Paid Medical Leave - Up to 12 weeks of 100% paid leave is provided for childbirth and medical needs.
  • InVacation - As a thank-you to our long-term team members, we offer a bonus after 7 years of service.
  • Wellness Subsidy - We provide a subsidy that can be applied toward gym memberships, fitness classes, and more.
  • Salary Range $152,000 - $228,000 USD plus bonus + equity
  • At Invoca, all new hires in the U.S. receive benefits starting on day one of employment.

Company info

  • Senior ML Engineer
  • About Invoca
  • Invoca is an AI-powered revenue execution platform that brings together marketing, commerce, and contact center teams to turn every customer interaction into measurable, profitable growth.
  • Join our dynamic, fast-growing team, where innovation and collaboration are at the core of our culture.
  • The Data Platform team owns the full ML lifecycle at Invoca, from model training and fine-tuning through inference optimization and production APIs.
  • Learn more on our blog or check out our open source projects.

Visa & Work Authorization

  • Please note that we are unable to provide initial visa sponsorship for this position.

This listing is sourced directly from Invoca's careers page and normalized into a canonical job model.