FAL
Staff Software Engineer, ML Performance & Systems
San Francisco · Staff+
Sponsorship not specified$180k-$250kDetected 61 days ago
Machine LearningPyTorch
About the role
- Work closely with our Applied ML team and customers (frontier labs on the media space) and make sure their workloads benefit from our accelerator.
Responsibilities
- Help fal maintain its frontier position on model performance for generative media models.
- Design and implement novel approaches to model serving architecture on top of our in-house inference engine, focusing on maximizing throughput while minimizing latency and resource usage.
- Develop performance monitoring and profiling tools to identify bottlenecks and optimization opportunities.
Requirements
- Strong foundation in systems programming with expertise in identifying and fixing bottlenecks.
- Proficient in Triton or willingness to learn with comparable experience in lower-level accelerator programming.
Skills
- fal is the generative media ecosystem powering the next generation of AI products.
Compensation
- $180,000 - $250,000 + equity + comprehensive benefits package
- We are currently hiring in downtown San Francisco.
Benefits
- Competitive salary and equity
- A lot of learning and growth opportunities
- Health, dental, and vision insurance (US)
Company info
- We are currently hiring in downtown San Francisco.
- We offer relocation assistance to San Francisco.
This listing is sourced directly from FAL's careers page and normalized into a canonical job model.