Deeproute.ai

Deeproute.ai

Member of Technical Staff (MTS) - Multimodal Foundation Models

Fremont, California, United States · Staff+

Sponsorship not specifiedDetected 47 days ago
Machine LearningPyTorchNLPComputer VisionRoboticsResearch

About the role

  • Focus Multimodal Foundation Models · Representation Learning · Method Innovation We are looking for strong technical builders and researchers who deeply understand foundation models and representation learning beyond simply applying existing frameworks. Ideal candidates should have: Strong experimental rigor Solid systems and modeling intuition Hands-on
  • engineering ability Interest in scalable multimodal AI systems for real-world autonomy We value people who can bridge research and production, and who care about robustness, scalability, efficiency, and practical deployment in large-scale autonomous driving systems. Responsibilities 1. Large-Scale Foundation Model Pretraining Develop scalable pretraining

Responsibilities

  • Large-Scale Foundation Model Pretraining Develop scalable pretraining pipelines for large-scale multimodal driving data

Requirements

  • Large-scale pretraining Hands-on experience with methods such as:

Nice to have

  • CVPR ICCV ECCV NeurIPS ICLR ICML

Benefits

  • Vision-language-action models Video foundation models
  • Vision Transformers (ViT) Video / temporal architectures Multimodal fusion and alignment Embedding and retrieval systems
  • Computer Vision
  • Machine Learning Robotics Computer Science Related fields Strong understanding of:
  • Foundation models Self-supervised learning Representation learning Multimodal learning
  • CLIP DINO / DINOv2 MAE Contrastive learning Masked modeling MoE or scalable transformer architectures

This listing is sourced directly from Deeproute.ai's careers page and normalized into a canonical job model.