xAI

xAI

Member of Technical Staff - RL Inference

Palo Alto, CA · Staff+

Sponsorship not specified$180k-$440kDetected 17 hours ago
PythonC++Distributed SystemsAlgorithmsPyTorchLLMsLeadershipCommunication

About the role

  • The RL infrastructure team is looking for an engineer to help with low precision RL training and inference.

Responsibilities

  • Design and optimize our inference stack for all shapes of RL workloads at SpaceXAI, from small scale ablations to production training runs.
  • Work closely with the modelling team to efficiently implement novel RL techniques and algorithms
  • Experience in building, debugging, and optimizing efficiency of large-scale distributed systems

Requirements

  • Experience in LLM inference
  • Proficiency in programming languages such as Python, C++ and/or Rust

Skills

  • Strong knowledge in quantization and numerics in LLM inference and training
  • Experience in developing inference engines, e.g. SGLang, vLLM
  • This organization is for individuals who appreciate challenging themselves and thrive on curiosity.
  • All employees are expected to be hands-on and to contribute directly to the company's mission.
  • Work ethic and strong prioritization skills are important.
  • All employees are expected to have strong communication skills.
  • They should be able to concisely and accurately share knowledge with their teammates.

Compensation

  • $180,000 - $440,000 USD

Equal opportunity

  • equal opportunity employer.

This listing is sourced directly from xAI's careers page and normalized into a canonical job model.