Exa
Software Engineer, Distributed Data Systems
San Francisco, California
Sponsorship not specifiedDetected 215 days ago
Vector DatabasesKafkaMachine LearningSparkData EngineeringHubSpotLogistics
About the role
- We now power search for Cursor, Cognition, HubSpot, and over 400,000 developers and have raised $350m from Lightspeed, Benchmark, and a16z.
- You'll have enormous autonomy in designing systems that scale to hundreds of petabytes.
- While we cannot guarantee your visa, we have historically been successful in sponsoring candidates from all over the world.
Responsibilities
- Experience building and operating large-scale distributed data processing pipelines
- An obsessive focus on reliability and building systems that don't page you at 3am
- Design a lakehouse architecture that handles 100+ PB of web crawl data
- Build streaming pipelines that process billions of documents per day for real-time indexing
Requirements
- Hands-on experience with streaming data systems (Kafka, Flink, or similar)
- Familiarity with Ray, Spark, or ClickHouse at production scale
- Experience with Lance or other vector-native storage formats
Benefits
- We offer premium healthcare benefits (medical, dental, vision), fertility benefits, 16 weeks of fully paid parental leave for all new parents, and a monthly wellness stipend to all of our employees.
Equal opportunity
- Exa is an equal opportunity employer.
Visa & Work Authorization
- We're happy to sponsor international candidates (e.g., STEM OPT, OPT, H1B, O1, E3).
- While we cannot guarantee your visa, we have historically been successful in sponsoring candidates from all over the world.
- If you receive an offer, our team will work hard to get you a visa.
This listing is sourced directly from Exa's careers page and normalized into a canonical job model.