Figure
Helix AI Engineer, Training Infrastructure
San Jose, CA · Full-time
Sponsorship not specified$200k-$400kDetected 8 days ago
PythonAlgorithmsAWSGCPAzureCloud PlatformsKubernetesTerraformAnsibleCI/CDDeep LearningPyTorchA/B TestingRobotics
About the role
- The goal of the company is to ship humanoid robots with human level intelligence.
- Figure is headquartered in San Jose, CA.
- Figure's vision is to deploy autonomous humanoids at a global scale.
Responsibilities
- Design, deploy, and maintain Figure's training clusters
- Work together with AI researchers to implement training of new model architectures at a large scale
- Implement distributed training, advanced parallelization strategies, and high-performance data loaders to reduce model development cycles
- Implement tooling for data processing, model experimentation, and continuous integration
- Minimum of 4 years of professional, full-time experience building reliable backend systems and infrastructure
- Figure is an AI robotics company developing autonomous general-purpose humanoid robots.
- Its robots are engineered to perform a variety of tasks in the home and commercial markets.
Requirements
- Bachelor's or Master's degree in Computer Science, Robotics, Engineering, or a related field
- Extensive professional experience with Python and PyTorch
- Proven track record of scaling and running large-scale training experiments personally on 800+ GPUs
- Experience managing HPC clusters for deep neural network training
Nice to have
- Experience contributing to or maintaining open-source distributed training frameworks (Megatron-LM, DeepSpeed, TorchTitan)
- Experience managing cloud infrastructure (AWS, Azure, GCP)
- Experience with job scheduling / orchestration tools (SLURM, Kubernetes, LSF, etc.)
- Experience with configuration management tools (Ansible, Terraform, Puppet, Chef, etc.)
- The pay offered for this position may vary based on several individual factors, including job-related knowledge, skills, and experience.
- This information will be shared if an employment offer is extended.
Compensation
- The US base salary range for this full-time position is between $200,000 - $400,000 annually.
Benefits
- Architect, optimize, and maintain scalable deep learning frameworks for training on massive robot datasets
- Bonus Qualifications
This listing is sourced directly from Figure's careers page and normalized into a canonical job model.