Cerebras Systems
CoDesign & NextGen Performance Engineer
Headquarters/Sunnyvale Office
Sponsorship not specifiedDetected 7 days ago
C++AlgorithmsMachine LearningDeep LearningLLMsComplianceElectrical EngineeringProblem Solving
About the role
- This role focuses on characterizing, analyzing, and optimizing the performance of state-of-the-art AI models running on Cerebras' breakthrough hardware.
Responsibilities
- Bring up and optimize performance on new generations of the Cerebras WSE.
- Build performance models (kernel-level, end-to-end) to estimate the performance of state of the art and customer ML models.
- Optimize and debug our kernel micro code and compiler algorithms to elevate ML model inference speed, throughput and compute utilization on the Cerebras WSE.
- Develop tools and infrastructure to help visualize performance data collected from the Wafer Scale Engine and our compute cluster.
Requirements
- 3+ years of experience in a relevant domain (Computer Architecture, CPU/GPU Performance, Kernel Optimization, HPC).
Skills
- Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups.
Benefits
- Exposure to and understanding of low-level deep learning / LLM math.
Company info
- Build a breakthrough AI platform beyond the constraints of the GPU.
- Publish and open source their cutting-edge AI research.
- Work on one of the fastest AI supercomputers in the world.
- Enjoy job stability with startup vitality.
- Our simple, non-corporate work culture that respects individual beliefs.
Apply directly at Cerebras Systems →Create a free account for alerts like thisView Cerebras Systems immigration profile
This listing is sourced directly from Cerebras Systems's careers page and normalized into a canonical job model.