ElastixAI
AI Compiler and Performance Engineer
Seattle
Sponsorship not specifiedDetected 147 days ago
PythonData StructuresAlgorithmsMachine LearningDeep LearningTensorFlowPyTorchNLPLLMsLLMOpsFPGACommunicationProblem Solving
About the role
- If you're excited about pushing AI performance to physical limits, and about shaping the future of large-scale inference, we'd love to meet you.
- This is not a "typical" compiler role, it's a chance to rethink the entire AI compute stack.
- At ElastixAI, if improving inference efficiency requires inventing new quantization schemes, rethinking graph-level optimizations, or modifying the hardware ISA, we do it.
Responsibilities
- Design and implement IR transformations, graph optimizations, kernel lowering, and code generation for novel hardware architectures.
- Collaborate with ML researchers to co-design algorithmic optimizations that yield real end-to-end performance gains.
- Build performance models, profiling tools, and benchmarking frameworks to identify bottlenecks and guide design decisions.
- Prototype and validate improvements across the entire stack - from PyTorch/XLA-level passes to custom kernel implementations.
- the best performance comes from holistic co-design, where every layer, from model architecture to kernels to silicon, works in harmony.
- In this role, you will design how LLM operations decompose into highly efficient proprietary kernel primitives, optimize execution pipelines, and co-develop abstractions with our hardware and ML teams.
- Prior startup experience or background building first-generation systems.
Requirements
- Strong proficiency in one or more programming languages such as Python and C++.
- Strong understanding of one or more of the following:
- Deep expertise on ML frameworks (e.g., PyTorch, TensorFlow, JAX) and understanding of ML model deployment challenges.
- Ability to thrive in a fast-paced, dynamic startup environment.
Nice to have
- PhD in Computer Science, Software Engineering, or a related field.
- Experience with custom hardware accelerators for ML inference.
- Contributions to open-source compiler or ML systems projects.
Compensation
- Competitive compensation and startup equity package
Benefits
- The opportunity to work on challenging problems at the intersection of ML, software, and systems.
- Competitive compensation and startup equity package
- Comprehensive medical, dental, and vision coverage (100% paid by employer)
- Life insurance and AD&D
- Flexible Time Off (FTO)
- Paid parental leave
- Gym or fitness benefit
- Commuter benefit
- Investment in employee learning & development
Company info
- ElastixAI, Inc.
- ElastixAI is an early-stage startup on a mission to reinvent AI inference infrastructure from the ground up.
- We're building a next-generation inference platform that delivers unprecedented efficiency by tightly integrating machine learning, software stack, and custom hardware.
- Our philosophy is simple: the best performance comes from holistic co-design, where every layer, from model architecture to kernels to silicon, works in harmony.
- We are looking for a deeply technical AI Compiler & Performance Engineer who thrives at the intersection of ML, compilers, and hardware.
- You'll have end-to-end ownership to explore radically new ideas and make them real.
- Break down LLM and transformer workloads into fine-grained primitives tailored to our proprietary compute hardware.
Apply directly at ElastixAI →Create a free account for alerts like thisView ElastixAI immigration profile
This listing is sourced directly from ElastixAI's careers page and normalized into a canonical job model.