Luma AI

Luma AI

Software Engineer - Data Infrastructure

Redwood City, CA

Sponsorship not specifiedDetected 26 days ago
PythonDistributed SystemsMachine LearningSparkData EngineeringRoboticsResearchProblem Solving

Stay score

odds of building a lasting career here

36Risky
Cap-exempt (no lottery)0
Sponsors this role74
Entry-level history0
PERM / green-card track0
Lottery odds40
Fits your clock70

Thin sponsorship signal and lottery-bound. A low-probability bet with your clock running. Prioritize cap-exempt roles and proven entry-level sponsors first.

Lottery odds assume a STEM candidate.

Personalize to your clock →

Employer immigration record

from this employer's Department of Labor filings

Files H-1B transfers

16 transfer filings in the last year, covering 16 workers. Median labor-condition decision: 7 days. An employer that already files transfers is one that can take over an existing H-1B.

Sourced from Department of Labor LCA, PERM and prevailing-wage disclosure data. Employer matching is by name, so figures may be split across an employer's legal entities. Absence of a filing means none appears in our copy of the data, not that none exists.

Community outcomes

No reports yet — be the first to help the next applicant.

About the role

  • This role requires a strong foundation in distributed systems and data engineering, with an emphasis on supporting complex machine learning workflows rather than traditional product data infrastructure.

Responsibilities

  • Collaborate with ML researchers and product teams to ensure data systems meet evolving needs
  • Develop and optimize large-scale data pipelines and batch processing jobs
  • Support the evaluation and adoption of new programming languages and frameworks relevant to data infrastructure
  • Collaborate with research & engineering teams to help define and refine best practices for data infrastructure development

Requirements

  • Proficiency in Python (or similar languages with willingness to learn Python) and experience with large-scale, high-throughput data infrastructure
  • Familiarity with distributed computing frameworks (e.g., Ray, Spark, Beam)
  • Experience sourcing, integrating, and optimizing data from diverse and large datasets
  • Ability to design and optimize data pipelines for ML research and internal teams
  • Strong problem-solving skills and understanding of data engineering at scale
  • Collaborative, product-focused mindset; comfortable in fast-paced environments
  • Comfortable working in a fast-paced, product-focused environment with a strong execution mindset
  • Open to candidates across seniority levels, from mid-level individual contributors to senior engineers and managers.

Nice to have

  • Prior experience working with complex data infrastructure or AI/ML platforms highly desirable
  • Experience with open source data infrastructure projects is a plus
  • Experience working in the robotics industry preferred

Benefits

  • Build and maintain scalable data infrastructure for high-throughput machine learning workflows

This listing is sourced directly from Luma AI's careers page and normalized into a canonical job model.