Inference
Machine Learning Researcher
San Francisco
Sponsorship not specified$250k-$350kDetected 197 days ago
GitMachine LearningPyTorchNLPLLMsCustomer SupportResearchAdaptability
About the role
- Your work will span from exploring new architectures and learning methods to optimizing latency and efficiency, with the goal of delivering better models to customers.
- Your north star is pushing the frontier of what's possible in LLM post-training.
- You'll explore new techniques, run rigorous experiments, and when something works, help bring it into production with the help of your teammates.
Responsibilities
- Develop and improve our distillation pipeline for training high-quality models from frontier teachers
- Train models for clients and run evaluations to validate research findings in production settings
- Create robust benchmarks and evaluation frameworks that ensure custom models match or exceed frontier performance
- Collaborate with applied engineers to bring successful research into production systems
- Document findings and share knowledge with the team
Requirements
- 3+ years of experience training AI models using PyTorch
- Hands-on experience with post-training LLMs using SFT, RLHF, DPO, or other alignment techniques
- Experience with LLM-specific training frameworks (e.g., Hugging Face Transformers, DeepSpeed, Megatron, TRL, or similar)
- Track record of implementing ideas from recent ML papers
- Experience training on NVIDIA GPUs at scale
Nice to have
- Publications in ML venues
- Experience with model distillation or knowledge transfer
- Experience with LLM speed optimization techniques
- Experience with distributed training and infrastructure at scale
- Contributions to open-source ML projects
- You don't need to tick every box.
- Curiosity and the ability to learn quickly matter more.
Compensation
- We offer competitive compensation, equity in a high-growth startup, and comprehensive benefits.
- The base salary range for this role is $250,000 - $350,000, plus equity and benefits, depending on experience.
Benefits
- Familiarity with vision encoders, multimodal models, or other modalities
- Run experiments with new learning methods, including novel approaches to SFT, RLHF, DPO, and other post-training techniques
- Perform reinforcement learning research to improve model alignment and capability
Company info
- You will be responsible for conducting research into experimental models, training systems, and modalities to create novel products for our customers.
- This includes training models for customers and running evaluations as part of validating your research.
Equal opportunity
- Inference.net http://Inference.net is an equal opportunity employer. We welcome applicants from all backgrounds and don't discriminate based on race, color, religion, gender, sexual orientation, national origin, genetics, disability, age, or veteran status.
- If you're excited about pushing the boundaries of custom AI research, we'd love to hear from you. Please send your resume and GitHub to amar@inference.net and/or here on Ashby.
- Inference.net http://Inference.net is an equal opportunity employer.
- We welcome applicants from all backgrounds and don't discriminate based on race, color, religion, gender, sexual orientation, national origin, genetics, disability, age, or veteran status.
Apply directly at Inference →Create a free account for alerts like thisView Inference immigration profile
This listing is sourced directly from Inference's careers page and normalized into a canonical job model.