ElastixAI
Hardware Design Engineer, AI Inference Engine
Seattle
Sponsorship not specifiedDetected 153 days ago
Machine LearningLLMsVerilogElectrical EngineeringHardware DesignResearchCommunicationProblem Solving
About the role
- DESCRIPTION Location: Seattle, WA (Hybrid - 3 days/week in office) About ElastixAI: ElastixAI is an early-stage startup poised to revolutionize AI inference infrastructure.
- We believe in providing a customizable and optimal inference experience, much like tailoring a high-performance computing system to specific needs.
Responsibilities
- Contribute to the architectural definition, design, and implementation of a novel AI inference engine optimized for our specific ML workloads.
- Collaborate closely with ML engineers to understand and influence ML directions
- Partner with cloud engineers to ensure the inference engine architecture aligns with target cloud hardware capabilities, deployment strategies, and performance/cost objectives.
- Stay at the forefront of AI accelerator research, identifying emerging techniques and technologies relevant to our co-design approach.
- Contribute to the RTL design, simulation, and verification efforts for the inference engine components.
- Drive the hardware roadmap for the inference engine, anticipating future AI model trends and optimization opportunities.
Requirements
- Experience with hardware implementation of ML optimization techniques (e.g., sparsity, quantization, pruning).
- Proficiency in Verilog or SystemVerilog for RTL design and simulation.
- Strong understanding of memory system architecture, on-chip interconnects, parallel processing, and distributed computing.
- Excellent problem-solving skills and the ability to analyze complex systems.
- Exceptional communication and interpersonal skills, with a demonstrated ability to work effectively in a highly interdisciplinary environment, collaborating with ML, software, and cloud/systems engineers.
- Ability to thrive in a fast-paced, dynamic startup environment with a strong bias for action/execution
Nice to have
- Knowledge of compiler technologies for AI models (e.g., MLIR, TVM).
- Familiarity with performance modeling and analysis tools.
- Experience with system-level integration and debugging.
- Contributions to relevant research publications or open-source projects.
- Understanding of cloud computing environments and deploying hardware accelerators in the cloud.
- High-speed inter-chip networking experience
Skills
- ElastixAI is an early-stage startup poised to revolutionize AI inference infrastructure.
Compensation
- Competitive compensation and startup equity package
Benefits
- The opportunity to work on challenging problems at the intersection of ML, software, and systems.
- Competitive compensation and startup equity package
- Comprehensive medical, dental, and vision coverage (100% paid by employer)
- Flexible Time Off (FTO)
- Paid parental leave
- Gym or fitness benefit
- Commuter benefit
- Investment in employee learning & development
Company info
- We are building the next-gen AI inference platform.
Apply directly at ElastixAI →Create a free account for alerts like thisView ElastixAI immigration profile
This listing is sourced directly from ElastixAI's careers page and normalized into a canonical job model.