Deeproute.ai
Member of Technical Staff (MTS) - Multimodal Foundation Models
Fremont, California, United States · Staff+
Sponsorship not specifiedDetected 47 days ago
Machine LearningPyTorchNLPComputer VisionRoboticsResearch
About the role
- Focus Multimodal Foundation Models · Representation Learning · Method Innovation We are looking for strong technical builders and researchers who deeply understand foundation models and representation learning beyond simply applying existing frameworks. Ideal candidates should have: Strong experimental rigor Solid systems and modeling intuition Hands-on
- engineering ability Interest in scalable multimodal AI systems for real-world autonomy We value people who can bridge research and production, and who care about robustness, scalability, efficiency, and practical deployment in large-scale autonomous driving systems. Responsibilities 1. Large-Scale Foundation Model Pretraining Develop scalable pretraining
Responsibilities
- Large-Scale Foundation Model Pretraining Develop scalable pretraining pipelines for large-scale multimodal driving data
Requirements
- Large-scale pretraining Hands-on experience with methods such as:
Nice to have
- CVPR ICCV ECCV NeurIPS ICLR ICML
Benefits
- Vision-language-action models Video foundation models
- Vision Transformers (ViT) Video / temporal architectures Multimodal fusion and alignment Embedding and retrieval systems
- Computer Vision
- Machine Learning Robotics Computer Science Related fields Strong understanding of:
- Foundation models Self-supervised learning Representation learning Multimodal learning
- CLIP DINO / DINOv2 MAE Contrastive learning Masked modeling MoE or scalable transformer architectures
Apply directly at Deeproute.ai →Create a free account for alerts like thisView Deeproute.ai immigration profile
This listing is sourced directly from Deeproute.ai's careers page and normalized into a canonical job model.