xAI
Member of Technical Staff - RL Inference
Palo Alto, CA · Staff+
Sponsorship not specified$180k-$440kDetected 17 hours ago
PythonC++Distributed SystemsAlgorithmsPyTorchLLMsLeadershipCommunication
About the role
- The RL infrastructure team is looking for an engineer to help with low precision RL training and inference.
Responsibilities
- Design and optimize our inference stack for all shapes of RL workloads at SpaceXAI, from small scale ablations to production training runs.
- Work closely with the modelling team to efficiently implement novel RL techniques and algorithms
- Experience in building, debugging, and optimizing efficiency of large-scale distributed systems
Requirements
- Experience in LLM inference
- Proficiency in programming languages such as Python, C++ and/or Rust
Skills
- Strong knowledge in quantization and numerics in LLM inference and training
- Experience in developing inference engines, e.g. SGLang, vLLM
- This organization is for individuals who appreciate challenging themselves and thrive on curiosity.
- All employees are expected to be hands-on and to contribute directly to the company's mission.
- Work ethic and strong prioritization skills are important.
- All employees are expected to have strong communication skills.
- They should be able to concisely and accurately share knowledge with their teammates.
Compensation
- $180,000 - $440,000 USD
Equal opportunity
- equal opportunity employer.
This listing is sourced directly from xAI's careers page and normalized into a canonical job model.