DDN
Senior/Staff AI Engineer
Remote - California · Staff+
Sponsorship not specifiedDetected 5 days ago
Distributed SystemsMachine LearningLLMsRAGMLOps
About the role
- - Build and optimize LLM serving and inference systems for production environments
- - Improve performance across GPU and CPU pathways
Responsibilities
- Build and optimize LLM serving and inference systems for production environments
- Design and scale systems that support RAG and retrieval-heavy AI workloads
- An engineer who has spent meaningful time building or optimizing production AI systems, not just experimenting with models
Requirements
- Deep hands-on experience working close to the systems layer - for example, improving how workloads run across GPU and CPU resources, reducing bottlenecks, or tuning infrastructure for better throughput and latency
- The ability to move comfortably between architecture decisions and hands-on implementation, especially in environments where efficiency and scale matter
- A background that suggests you can operate in technically demanding environments, whether that comes from AI infrastructure, high-performance systems, storage platforms, or adjacent distributed systems work
Nice to have
- PhD preferred, but far less important than having built serious systems in the real world
This listing is sourced directly from DDN's careers page and normalized into a canonical job model.