Rhoda AI

Rhoda AI

Research Member of Technical Staff- Video Generation Modeling

Mountain View · Staff+

Sponsorship not specifiedDetected 64 days ago
Node.jsMachine LearningPyTorchNLPLLMsRoboticsHardware DesignResearchCollaboration

About the role

  • We've raised over $450M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality.
  • Our approach formulates robot control as video prediction - we pre-train causal video generation models on web-scale video data, then adapt them to predict robot actions from real-world demonstrations.
  • You'll work on the core architectures, training objectives, and scaling strategies that determine how well our models learn from internet-scale video.

Responsibilities

  • Design and train large-scale causal video generation models on web-scale video data
  • Develop and validate training objectives, model architectures, and data mixtures for video prediction at scale
  • Build systematic evaluations to measure video generation quality, long-horizon prediction fidelity, and downstream robot task performance
  • Run rigorous ablations and benchmarking to understand what drives model quality at scale
  • Collaborate closely with data & evaluation, post-training, and training systems teams to translate research ideas into working systems
  • We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots.
  • Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design.

Requirements

  • Strong research taste: ability to identify high-leverage questions and cut through noise

Nice to have

  • Comfort operating in a fast-moving, ambiguous startup environment
  • senior/MTS candidates execute complex projects with strong fundamentals and growing scope
  • PhD in ML, CS, Robotics, or a related field - or equivalent research/industry experience
  • Strong publication record at NeurIPS, ICML, ICLR, CVPR, CoRL, etc. (especially valued for RS track)
  • Prior work specifically on video generation models (autoregressive video, diffusion transformers, world models, or causal video architectures)
  • Experience with large-scale autoregressive language model pretraining and scaling
  • Familiarity with web-scale video datasets and video data curation pipelines
  • Familiarity with distributed training and multi-node infrastructure

Company info

  • What We're Looking For
  • At Rhoda AI, we're building the next generation of generalist intelligent robots.
  • We're looking for Research Scientists and Research Engineers to push the frontier of large-scale pre-training for our video action model.

This listing is sourced directly from Rhoda AI's careers page and normalized into a canonical job model.