ElastixAI

ElastixAI

AI Compiler and Performance Engineer

Seattle

Sponsorship not specifiedDetected 147 days ago
PythonData StructuresAlgorithmsMachine LearningDeep LearningTensorFlowPyTorchNLPLLMsLLMOpsFPGACommunicationProblem Solving

About the role

  • If you're excited about pushing AI performance to physical limits, and about shaping the future of large-scale inference, we'd love to meet you.
  • This is not a "typical" compiler role, it's a chance to rethink the entire AI compute stack.
  • At ElastixAI, if improving inference efficiency requires inventing new quantization schemes, rethinking graph-level optimizations, or modifying the hardware ISA, we do it.

Responsibilities

  • Design and implement IR transformations, graph optimizations, kernel lowering, and code generation for novel hardware architectures.
  • Collaborate with ML researchers to co-design algorithmic optimizations that yield real end-to-end performance gains.
  • Build performance models, profiling tools, and benchmarking frameworks to identify bottlenecks and guide design decisions.
  • Prototype and validate improvements across the entire stack - from PyTorch/XLA-level passes to custom kernel implementations.
  • the best performance comes from holistic co-design, where every layer, from model architecture to kernels to silicon, works in harmony.
  • In this role, you will design how LLM operations decompose into highly efficient proprietary kernel primitives, optimize execution pipelines, and co-develop abstractions with our hardware and ML teams.
  • Prior startup experience or background building first-generation systems.

Requirements

  • Strong proficiency in one or more programming languages such as Python and C++.
  • Strong understanding of one or more of the following:
  • Deep expertise on ML frameworks (e.g., PyTorch, TensorFlow, JAX) and understanding of ML model deployment challenges.
  • Ability to thrive in a fast-paced, dynamic startup environment.

Nice to have

  • PhD in Computer Science, Software Engineering, or a related field.
  • Experience with custom hardware accelerators for ML inference.
  • Contributions to open-source compiler or ML systems projects.

Compensation

  • Competitive compensation and startup equity package

Benefits

  • The opportunity to work on challenging problems at the intersection of ML, software, and systems.
  • Competitive compensation and startup equity package
  • Comprehensive medical, dental, and vision coverage (100% paid by employer)
  • Life insurance and AD&D
  • Flexible Time Off (FTO)
  • Paid parental leave
  • Gym or fitness benefit
  • Commuter benefit
  • Investment in employee learning & development

Company info

  • ElastixAI, Inc.
  • ElastixAI is an early-stage startup on a mission to reinvent AI inference infrastructure from the ground up.
  • We're building a next-generation inference platform that delivers unprecedented efficiency by tightly integrating machine learning, software stack, and custom hardware.
  • Our philosophy is simple: the best performance comes from holistic co-design, where every layer, from model architecture to kernels to silicon, works in harmony.
  • We are looking for a deeply technical AI Compiler & Performance Engineer who thrives at the intersection of ML, compilers, and hardware.
  • You'll have end-to-end ownership to explore radically new ideas and make them real.
  • Break down LLM and transformer workloads into fine-grained primitives tailored to our proprietary compute hardware.

This listing is sourced directly from ElastixAI's careers page and normalized into a canonical job model.