Exa

Exa

Software Engineer, Distributed Data Systems

San Francisco, California

Sponsorship not specifiedDetected 215 days ago
Vector DatabasesKafkaMachine LearningSparkData EngineeringHubSpotLogistics

About the role

  • We now power search for Cursor, Cognition, HubSpot, and over 400,000 developers and have raised $350m from Lightspeed, Benchmark, and a16z.
  • You'll have enormous autonomy in designing systems that scale to hundreds of petabytes.
  • While we cannot guarantee your visa, we have historically been successful in sponsoring candidates from all over the world.

Responsibilities

  • Experience building and operating large-scale distributed data processing pipelines
  • An obsessive focus on reliability and building systems that don't page you at 3am
  • Design a lakehouse architecture that handles 100+ PB of web crawl data
  • Build streaming pipelines that process billions of documents per day for real-time indexing

Requirements

  • Hands-on experience with streaming data systems (Kafka, Flink, or similar)
  • Familiarity with Ray, Spark, or ClickHouse at production scale
  • Experience with Lance or other vector-native storage formats

Benefits

  • We offer premium healthcare benefits (medical, dental, vision), fertility benefits, 16 weeks of fully paid parental leave for all new parents, and a monthly wellness stipend to all of our employees.

Equal opportunity

  • Exa is an equal opportunity employer.

Visa & Work Authorization

  • We're happy to sponsor international candidates (e.g., STEM OPT, OPT, H1B, O1, E3).
  • While we cannot guarantee your visa, we have historically been successful in sponsoring candidates from all over the world.
  • If you receive an offer, our team will work hard to get you a visa.

This listing is sourced directly from Exa's careers page and normalized into a canonical job model.