FAL

FAL

Staff Software Engineer, ML Performance & Systems

San Francisco · Staff+

Sponsorship not specified$180k-$250kDetected 61 days ago
Machine LearningPyTorch

About the role

  • Work closely with our Applied ML team and customers (frontier labs on the media space) and make sure their workloads benefit from our accelerator.

Responsibilities

  • Help fal maintain its frontier position on model performance for generative media models.
  • Design and implement novel approaches to model serving architecture on top of our in-house inference engine, focusing on maximizing throughput while minimizing latency and resource usage.
  • Develop performance monitoring and profiling tools to identify bottlenecks and optimization opportunities.

Requirements

  • Strong foundation in systems programming with expertise in identifying and fixing bottlenecks.
  • Proficient in Triton or willingness to learn with comparable experience in lower-level accelerator programming.

Skills

  • fal is the generative media ecosystem powering the next generation of AI products.

Compensation

  • $180,000 - $250,000 + equity + comprehensive benefits package
  • We are currently hiring in downtown San Francisco.

Benefits

  • Competitive salary and equity
  • A lot of learning and growth opportunities
  • Health, dental, and vision insurance (US)

Company info

  • We are currently hiring in downtown San Francisco.
  • We offer relocation assistance to San Francisco.

This listing is sourced directly from FAL's careers page and normalized into a canonical job model.