Cerebras Systems

Cerebras Systems

CoDesign & NextGen Performance Engineer

Headquarters/Sunnyvale Office

Sponsorship not specifiedDetected 7 days ago
C++AlgorithmsMachine LearningDeep LearningLLMsComplianceElectrical EngineeringProblem Solving

About the role

  • This role focuses on characterizing, analyzing, and optimizing the performance of state-of-the-art AI models running on Cerebras' breakthrough hardware.

Responsibilities

  • Bring up and optimize performance on new generations of the Cerebras WSE.
  • Build performance models (kernel-level, end-to-end) to estimate the performance of state of the art and customer ML models.
  • Optimize and debug our kernel micro code and compiler algorithms to elevate ML model inference speed, throughput and compute utilization on the Cerebras WSE.
  • Develop tools and infrastructure to help visualize performance data collected from the Wafer Scale Engine and our compute cluster.

Requirements

  • 3+ years of experience in a relevant domain (Computer Architecture, CPU/GPU Performance, Kernel Optimization, HPC).

Skills

  • Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups.

Benefits

  • Exposure to and understanding of low-level deep learning / LLM math.

Company info

  • Build a breakthrough AI platform beyond the constraints of the GPU.
  • Publish and open source their cutting-edge AI research.
  • Work on one of the fastest AI supercomputers in the world.
  • Enjoy job stability with startup vitality.
  • Our simple, non-corporate work culture that respects individual beliefs.

This listing is sourced directly from Cerebras Systems's careers page and normalized into a canonical job model.