Nuancelabs

Nuancelabs

Member of Technical Staff — Pretraining Infra (Experienced)

Seattle, Washington · Staff+

H1B sponsorship available$300k-$400kDetected 40 days ago
ReactDistributed SystemsMachine LearningPyTorchLLMsLogisticsHRISFirewallResearchCommunicationAdaptability

About the role

  • Direct impact on what models we can train, how fast research can iterate, and how reliably we scale.
  • Evolve the platform as model architectures, training recipes, data mixtures, sequence lengths, hardware constraints, and research directions change.
  • Our models are omni from the ground up (audio, video, language, real-time full-duplex), which introduces systems challenges beyond standard

Responsibilities

  • Own the distributed training stack for omni model pretraining, from 0→1 system design to 1→10 scaling across large GPU clusters.
  • Optimize large-scale training performance across parallelism strategy, GPU communication, memory usage, data throughput, MFU, step time, and end-to-end training efficiency.
  • Build and operate the core training runtime: job orchestration, distributed execution, checkpointing, recovery, monitoring, and debugging for long-running training jobs.
  • Build infrastructure for omni training workloads: high-throughput audio/video/text data loading, temporal alignment, variable sequence handling, multimodal synchronization, and memory-efficient training.
  • We believe diverse teams build better AI.

Requirements

  • Practical experience with at least one major large-scale training stack such as Megatron, PyTorch FSDP, DeepSpeed, or equivalent internal infrastructure.
  • Experience training large omni or multimodal models involving audio, video, text, or long-context temporal data.
  • Experience with adjacent infrastructure areas such as RL/post-training, data infrastructure, synthetic data generation, evaluation, or serving.

Nice to have

  • Hands-on experience running large-scale distributed training jobs across large GPU clusters
  • experience at hundreds of GPUs minimum, 1,000+ GPUs a strong plus.

Skills

  • Do your best work with the best tools, including unlimited tokens.

Compensation

  • $300,000 - $400,000 base salary, plus meaningful equity. We think long-term ownership matters and structure equity accordingly.
  • Visa sponsorship: We sponsor visas (O-1, H-1B, green card) from day one.
  • AI-native tooling: Do your best work with the best tools, including unlimited tokens.

Benefits

  • Health: HSA plan with ~$2,000 in annual company contributions - roughly 2x what most big tech companies put in.
  • Time off: 15 days of PTO plus public holidays, and we close the office for a full week at year-end.
  • Commuter benefits: We help cover the cost of getting to the office.

Company info

  • About Nuance Labs
  • Labs is building photorealistic, real-time AI avatars with emotional intelligence:
  • a full-duplex audiovisual system that can listen, speak, react, interrupt, and respond like a real person.
  • We're a research company, with PhDs from MIT, UW, Oxford, CMU, and Johns Hopkins, and industry experience from Apple, Meta, Amazon AGI, and Discord.
  • The team is small, the work is real, and the problems are unsolved.
  • How Nuance Differentiates
  • Most conversational AI avatars today are hacks - a face slapped on a speech-to-speech pipeline, stuck in the uncanny valley: emotionless, mechanical, one-turn-at-a-time.
  • Current systems take 2-5 seconds to respond; natural conversation requires sub-500ms.
  • That's a 10x improvement, and it demands rethinking the entire stack.
  • We're looking for a deeply technical MTS to own distributed training infrastructure for large-scale omni model pretraining.

Equal opportunity

  • equal opportunity employer.

Visa & Work Authorization

  • We sponsor visas (O-1, H-1B, green card) from day one.

This listing is sourced directly from Nuancelabs's careers page and normalized into a canonical job model.