Luma AI

Luma AI

Software Engineer, Inference

Redwood City, CA

Sponsorship not specified$30k-$60kDetected 43 days ago
PythonRedisDockerKubernetesCI/CDLinuxMachine LearningPyTorchLLMsSystems EngineeringResearch

Stay score

odds of building a lasting career here

35Risky
Cap-exempt (no lottery)0
Sponsors this role74
Entry-level history0
PERM / green-card track0
Lottery odds (Level I)39
Fits your clock70

Thin sponsorship signal and lottery-bound (~15% per draw). A low-probability bet with your clock running. Prioritize cap-exempt roles and proven entry-level sponsors first.

Lottery odds assume a STEM candidate.

Personalize to your clock →

H-1B wage level

the lottery is wage-weighted — each level is one more entry

Level I · 1×
Level I$137,0931 entry
Level II$165,1312 entries
Level III$193,1493 entries
Level IV$221,1874 entries

The bottom of this range ($30,000) is below the Level I prevailing wage of $137,093. An H-1B cannot be filed below the prevailing wage, so an offer at the floor of this band could not be sponsored as posted.

DOL prevailing wage, 2026-27 wage year · Software Developers (15-1252) · San Francisco-Oakland-Fremont, CA. Wage level is derived by USCIS from the offered wage, occupation and worksite; the occupation shown is inferred from the job title.

Employer immigration record

from this employer's Department of Labor filings

Files H-1B transfers

16 transfer filings in the last year, covering 16 workers. Median labor-condition decision: 7 days. An employer that already files transfers is one that can take over an existing H-1B.

Sourced from Department of Labor LCA, PERM and prevailing-wage disclosure data. Employer matching is by name, so figures may be split across an employer's legal entities. Absence of a filing means none appears in our copy of the data, not that none exists.

Community outcomes

No reports yet — be the first to help the next applicant.

About the role

  • This is large-scale inference systems work: scheduling, fleet management, deployment pipelines, and reliability across clusters and hardware providers.
  • It fits a strong systems engineer comfortable with model serving and Kubernetes at scale.
  • If you want pure modeling rather than the systems that run models, this is firmly the systems side.

Requirements

  • Experience deploying models with PyTorch, Hugging Face, vLLM, SGLang, TensorRT-LLM, or similar.
  • Experience with queues, scheduling, traffic control, and fleet management at scale.
  • Experience with Linux, Docker, and Kubernetes, and with orchestration, deployment, and scheduling.
  • Familiarity with Redis and S3-compatible storage.
  • Strong Python and system-architecture skills.

Nice to have

  • Modern networking stacks including RDMA (RoCE, InfiniBand, NVLink).
  • High-performance large-scale ML systems (100+ GPUs).
  • CUDA, and FFmpeg or multimedia processing.
  • Luma is an equal opportunity employer.

Compensation

  • $30k-$60k

Benefits

  • We believe multimodality is critical for intelligence - the next step beyond language models comes from vision.

Company info

  • Luma's mission is to build unified general intelligence that can generate, understand, and operate in the physical world.

This listing is sourced directly from Luma AI's careers page and normalized into a canonical job model.