Sequen AI

Sequen AI

Staff, MLOps Engineer

United States · Staff+

Sponsorship not specified$220k-$280kDetected 14 days ago
PythonRustDistributed SystemsAWSGCPAzureCloud PlatformsDockerKubernetesCI/CDMachine LearningData EngineeringLLMsMLOpsA/B TestingResearchCollaborationPipeline Integrity

About the role

  • This is a foundational, purely infrastructure-focused role sitting at the intersection of machine learning, backend distributed systems, and platform performance.
  • instead, your primary customer will be our internal ML research scientists.
  • Your mission is to make model serving, evaluation, and scaling completely seamless, reliable, and highly optimized in high-throughput production environments.

Responsibilities

  • Build ML infrastructure: Design, operate, and maintain robust systems for low-latency model deployment, distributed inference pipelines, and automated real-time telemetry.
  • Implement model CI/CD: Build reliable infrastructure for automated model versioning, canary releases, hot-swappable container rollouts, and zero-downtime rollbacks.
  • Drive system observability: Architect and monitor real-time pipelines to track model performance, data distribution drift, and system reliability anomalies.
  • Develop evaluation loops: Engineer robust evaluation pipelines and feedback loops to continuously validate live inference accuracy and prevent training-serving skew.
  • Optimize platform bottlenecks: Proactively isolate and eliminate performance bottlenecks across our serving layers, improving core tooling, model warm-up times, and researcher velocity.
  • Collaborate with research: Partner closely with our internal ML researchers and backend engineers to translate experimental model breakthroughs into resilient, production-grade serving topologies.
  • Pioneering systems: The unique opportunity to build and scale category-defining, low-latency ML platforms backed by proven, highly quantified customer revenue results.

Requirements

  • Maintain deep, production-grade proficiency with Python and PyTorch.
  • Bring production experience or active, hands-on familiarity with Rust for low-overhead systems engineering.
  • Familiarity with enterprise-grade feature stores, advanced experiment tracking, and systematic model evaluation frameworks.
  • Core ML framework mastery: Maintain deep, production-grade proficiency with Python and PyTorch.
  • Rust systems proficiency: Bring production experience or active, hands-on familiarity with Rust for low-overhead systems engineering.

Compensation

  • Top-tier reward: Highly competitive base salary, uncapped performance metrics, and meaningful early-employee equity.

Benefits

  • Top-tier reward: Highly competitive base salary, uncapped performance metrics, and meaningful early-employee equity.
  • Premium benefits: Full premium medical/dental/vision coverage, unlimited paid time off, and a highly collaborative, world-class engineering culture.
  • $220,000 - $280,000 Base + Performance Bonus + Meaningful Equity
  • Proven track record: Bring 4-8+ years of practical experience in MLOps, Machine Learning Engineering, or distributed platform/infrastructure engineering.
  • Low-latency serving expertise: Demonstrate hands-on experience deploying and serving ultra-low-latency machine learning models under heavy, real-time concurrent workloads.
  • Systems core maturity: Bring a solid, first-principles understanding of the complete machine learning lifecycle, asynchronous event-driven patterns, and distributed systems.

Company info

  • Sequen provides an integrated platform that pairs cutting-edge frontier ranking models with the infrastructure to run them in production-at sub-10ms latency and enterprise scale.
  • We are a small, highly technical, early-stage team focused on turning recent advances in AI into production-grade systems that operate under unforgiving real-world constraints.
  • The problems we work on are deeply open-ended, where minor optimizations in algorithmic multi-stage retrieval and model routing translate directly into millions of dollars in client revenue.
  • The world's largest retailers, marketplaces, and travel platforms use Sequen to rank, recommend, and personalize, with an autonomous research engine that compounds model performance into revenue and margin lift measured in hundreds of millions of dollars per customer.
  • A foundational, high-autonomy role directly shaping the core deployment and serving topology of a category-defining AI infrastructure company.

This listing is sourced directly from Sequen AI's careers page and normalized into a canonical job model.