DDN

DDN

Senior/Staff AI Engineer

Remote - California · Staff+

Sponsorship not specifiedDetected 5 days ago
Distributed SystemsMachine LearningLLMsRAGMLOps

About the role

  • - Build and optimize LLM serving and inference systems for production environments
  • - Improve performance across GPU and CPU pathways

Responsibilities

  • Build and optimize LLM serving and inference systems for production environments
  • Design and scale systems that support RAG and retrieval-heavy AI workloads
  • An engineer who has spent meaningful time building or optimizing production AI systems, not just experimenting with models

Requirements

  • Deep hands-on experience working close to the systems layer - for example, improving how workloads run across GPU and CPU resources, reducing bottlenecks, or tuning infrastructure for better throughput and latency
  • The ability to move comfortably between architecture decisions and hands-on implementation, especially in environments where efficiency and scale matter
  • A background that suggests you can operate in technically demanding environments, whether that comes from AI infrastructure, high-performance systems, storage platforms, or adjacent distributed systems work

Nice to have

  • PhD preferred, but far less important than having built serious systems in the real world

This listing is sourced directly from DDN's careers page and normalized into a canonical job model.