Lambda

Lambda

Staff Software Engineer - Infrastructure Storage

San Francisco Office (Fremont St) · Staff+

Sponsorship not specifiedDetected 5 days ago
RustC++Distributed SystemsCloud PlatformsKubernetesPrometheusGrafanaMachine LearningRoadmappingSystems EngineeringResearchLeadershipCollaborationMentoring

About the role

  • We are seeking a seasoned Staff Storage Software Engineer with deep experience designing and deploying storage protocol solutions at scale across object, block, and file paradigms.
  • This is a unique opportunity to work at the intersection of large-scale distributed systems and the rapidly evolving field of artificial intelligence infrastructure.
  • This is an opportunity to have a significant impact on the future of AI.

Responsibilities

  • Technical Leadership: Set technical direction for storage software architecture across petabyte-scale deployments, authoring and reviewing design docs, mentoring senior engineers, and serving as the technical anchor for cross-functional initiatives spanning storage, networking, compute, and control plane teams.
  • Execution: Design, develop, and maintain high-performance storage systems software across file (NFS, SMB, Lustre), block (NVMe-oF, iSCSI), and object (S3) protocols.
  • Build distributed systems for orchestrating storage resources, integrate with NVMe/GPU-direct/DPU-accelerated hardware, and troubleshoot complex production issues across performance, protocol, and hardware failure domains.
  • Own the full lifecycle from requirements and design through deployment, monitoring, and maintenance, including benchmarking, profiling, and capacity planning tooling.
  • Collaboration: Partner closely with storage software, networking, control plane, Kubernetes, observability, compute, and fleet engineering teams to deliver cross-functional infrastructure initiatives, define and track storage SLOs/SLIs, and ensure reliable deployment and maintenance of distributed storage infrastructure.
  • Innovate: Stay current with AI and HPC storage research, evaluate emerging protocols and hardware (from open-source filesystems to vendor-specific accelerated storage), and optimize solutions for AI workloads including checkpoint I/O, high-throughput dataset serving, and latency-sensitive inference pipelines.
  • Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove

Nice to have

  • Strong proficiency in C, C++, Rust, or Go.
  • Ability to write high-performance, concurrent, production-grade systems code.
  • Familiarity with DPDK/SPDK and kernel-bypass data paths is a plus
  • kernel-level storage driver or storage daemon experience is even better.
  • Systems-Level Programming: Strong proficiency in C, C++, Rust, or Go.

Skills

  • Working knowledge of NVMe, NVMe-oF, RDMA (RoCE or InfiniBand), and DPUs (e.g., NVIDIA BlueField).
  • Nice to Have
  • Experience with NVIDIA BlueField DPUs or SuperNICs for accelerated storage data paths, including GPUDirect Storage implementation.
  • Deep production experience with enterprise or HPC storage platforms: Vast Data, Weka, NetApp, or Lustre.
  • Experience deploying and operating Ceph like service at scale (100PB+) in an HPC or AI infrastructure environment.
  • Familiarity with emerging storage technologies such as CXL memory pooling, computational storage, or ZNS (Zoned Namespace) SSDs.
  • Experience contributing to or maintaining open-source storage projects (e.g., Ceph, DAOS, Lustre, MinIO).
  • Salary Range Information
  • The annual salary range for this position has been set based on market data and other factors.
  • About Lambda

Compensation

  • The annual salary range for this position has been set based on market data and other factors.
  • However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.
  • We offer generous cash & equity compensation

Benefits

  • We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG
  • Health, dental, and vision coverage for you and your dependents
  • Wellness and commuter stipends for select roles
  • Flexible paid time off plan that we all actually use

Company info

  • We offer generous cash & equity compensation
  • 401k Plan with 2% company match (USA employees)

Equal opportunity

  • Lambda is an Equal Opportunity employer.
  • Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
  • Equal Opportunity Employer

Visa & Work Authorization

  • ational origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law

This listing is sourced directly from Lambda's careers page and normalized into a canonical job model.