Nuancelabs
Member of Technical Staff — Pretraining Infra (Experienced)
Seattle, Washington · Staff+
H1B sponsorship available$300k-$400kDetected 40 days ago
ReactDistributed SystemsMachine LearningPyTorchLLMsLogisticsHRISFirewallResearchCommunicationAdaptability
About the role
- Direct impact on what models we can train, how fast research can iterate, and how reliably we scale.
- Evolve the platform as model architectures, training recipes, data mixtures, sequence lengths, hardware constraints, and research directions change.
- Our models are omni from the ground up (audio, video, language, real-time full-duplex), which introduces systems challenges beyond standard
Responsibilities
- Own the distributed training stack for omni model pretraining, from 0→1 system design to 1→10 scaling across large GPU clusters.
- Optimize large-scale training performance across parallelism strategy, GPU communication, memory usage, data throughput, MFU, step time, and end-to-end training efficiency.
- Build and operate the core training runtime: job orchestration, distributed execution, checkpointing, recovery, monitoring, and debugging for long-running training jobs.
- Build infrastructure for omni training workloads: high-throughput audio/video/text data loading, temporal alignment, variable sequence handling, multimodal synchronization, and memory-efficient training.
- We believe diverse teams build better AI.
Requirements
- Practical experience with at least one major large-scale training stack such as Megatron, PyTorch FSDP, DeepSpeed, or equivalent internal infrastructure.
- Experience training large omni or multimodal models involving audio, video, text, or long-context temporal data.
- Experience with adjacent infrastructure areas such as RL/post-training, data infrastructure, synthetic data generation, evaluation, or serving.
Nice to have
- Hands-on experience running large-scale distributed training jobs across large GPU clusters
- experience at hundreds of GPUs minimum, 1,000+ GPUs a strong plus.
Skills
- Do your best work with the best tools, including unlimited tokens.
Compensation
- $300,000 - $400,000 base salary, plus meaningful equity. We think long-term ownership matters and structure equity accordingly.
- Visa sponsorship: We sponsor visas (O-1, H-1B, green card) from day one.
- AI-native tooling: Do your best work with the best tools, including unlimited tokens.
Benefits
- Health: HSA plan with ~$2,000 in annual company contributions - roughly 2x what most big tech companies put in.
- Time off: 15 days of PTO plus public holidays, and we close the office for a full week at year-end.
- Commuter benefits: We help cover the cost of getting to the office.
Company info
- About Nuance Labs
- Labs is building photorealistic, real-time AI avatars with emotional intelligence:
- a full-duplex audiovisual system that can listen, speak, react, interrupt, and respond like a real person.
- We're a research company, with PhDs from MIT, UW, Oxford, CMU, and Johns Hopkins, and industry experience from Apple, Meta, Amazon AGI, and Discord.
- The team is small, the work is real, and the problems are unsolved.
- How Nuance Differentiates
- Most conversational AI avatars today are hacks - a face slapped on a speech-to-speech pipeline, stuck in the uncanny valley: emotionless, mechanical, one-turn-at-a-time.
- Current systems take 2-5 seconds to respond; natural conversation requires sub-500ms.
- That's a 10x improvement, and it demands rethinking the entire stack.
- We're looking for a deeply technical MTS to own distributed training infrastructure for large-scale omni model pretraining.
Equal opportunity
- equal opportunity employer.
Visa & Work Authorization
- We sponsor visas (O-1, H-1B, green card) from day one.
Apply directly at Nuancelabs →Create a free account for alerts like thisView Nuancelabs immigration profile
This listing is sourced directly from Nuancelabs's careers page and normalized into a canonical job model.