Rhoda AI
Research Member of Technical Staff - Data & Evaluation
Mountain View · Staff+
Sponsorship not specifiedDetected 56 days ago
Machine LearningData EngineeringComputer VisionRoboticsHardware DesignResearch
About the role
- We've raised over $450M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality.
Responsibilities
- Design and implement scalable curation pipelines for web-scale video pretraining data: ingestion, deduplication, quality filtering, and content classification across internet-scale video corpora
- Develop video-specific annotation frameworks and quality filters - motion quality, scene diversity, action content, temporal coherence - to improve pretraining signal
- Build evaluation frameworks and benchmarks to measure causal video model capabilities: prediction quality, temporal coherence, long-horizon rollout fidelity, and downstream robot task performance
- Research and implement data selection, mixing, and weighting strategies that improve video generation quality and transfer to robotic control
- Collaborate closely with pre-training and post-training teams to ensure data quality and evaluation methodology drive research decisions
- We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots.
- Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design.
- This team owns web-scale video data curation, annotation pipelines, and evaluation methodology - directly determining the quality of the video pretraining distribution and how clearly we can measure model progress.
- Ability to design evaluations for video generation models that are diagnostic, reproducible, and actionable
- Staff-level candidates are expected to define technical direction and drive research strategy independently; senior/MTS candidates execute complex projects with strong fundamentals and growing scope
Requirements
- Strong understanding of data-centric ML and how web video data quality affects large generative model performance
- Familiarity with video-specific data characteristics: temporal structure, motion quality, scene diversity, and action content
- Solid ML fundamentals with hands-on experience training or evaluating large generative models
- Experience with large-scale web video dataset curation (e.g., WebVid, HowTo100M, Ego4D, or similar)
- Familiarity with video generation quality metrics (FVD, perceptual quality, motion consistency)
- Experience running VLM or CLIP-style inference at scale for automated video filtering and annotation
Nice to have
- Nice to Have (But Not Required)
Benefits
- PhD or strong research background in ML, computer vision, or a related field
- Deploy and scale vision-language models (VLMs) and video understanding models for automated annotation, filtering, and content scoring at web scale
Company info
- At Rhoda AI, we're building the next generation of generalist intelligent robots.
- We're looking for Research Scientists and Research Engineers to build the data and evaluation foundations for our video action model.
Apply directly at Rhoda AI →Create a free account for alerts like thisView Rhoda AI immigration profile
This listing is sourced directly from Rhoda AI's careers page and normalized into a canonical job model.