Kla
High Performance Compute (HPC) Software Engineer – HPC SW Systems
Ann Arbor, MI
Sponsorship not specified$106k-$180kDetected 30 days ago
JavaC++Node.jsAlgorithmsDockerKubernetesLinuxElectrical EngineeringAR/VRResearchCommunicationCollaboration
About the role
- Company Overview KLA is a global leader in diversified electronics for the semiconductor manufacturing ecosystem.
- No laptop, smartphone, wearable device, voice-controlled gadget, flexible screen, VR device or smart car would have made it into your hands without us.
Responsibilities
- Design, develop, and optimize HPC software running on large-scale Linux clusters, including distributed and parallel workloads (MPI, multithreading, GPU-accelerated pipelines, containerized workloads).
- Optimize application performance and power utilization across CPU, memory, storage, and network subsystem, with attention to throughput, latency, and scaling behavior.
- Collaborate with hardware and systems teams to define HPC node, storage, and interconnect requirements based on software and algorithm needs.
- Contribute to best practices, and design reviews for new platforms and refresh cycles.
Requirements
- Bachelor's or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent practical experience.
- Proficiency in Java and/or C++ and/or other system-level or performance-oriented languages.
- Doctorate (Academic) Degree and 0 years related work experience
- Master's Level Degree and related work experience of 3 years
- Bachelor's Level Degree and related work experience of 5 years
Nice to have
- Experience with containerized HPC environments (Docker, Singularity/Apptainer, Kubernetes in HPC contexts).
- Familiarity with high-speed interconnects, storage architectures, and performance benchmarking.
- Exposure to rack integration, including cabling, power distribution, cooling, and system bring-up.
- Experience in semiconductor, manufacturing, or high-reliability systems environments.
- Ability to reason about system reliability, MTBF/MTBA, and failure modes in large compute installations.
- What Makes This Role Unique at KLA
- Work on mission-critical HPC platforms that directly impact semiconductor manufacturing capability.
- See your work deployed at scale in real production tools-not just in the data center.
Skills
- CPUs, memory hierarchies, storage, networking (Ethernet / InfiniBand).
- Practical experience working with clusters, servers, or rack-scale systems in lab or production environments.
- Strong debugging skills across software, OS, and hardware boundaries.
- Virtually every electronic device in the world is produced using our technologies.
Compensation
- $105,900.00 - $180,000.00 Annually
- Our pay ranges are determined by role, level, and location.
- The range displayed reflects the pay for this position in the primary location identified in this posting.
- Actual pay depends on several factors, including state minimum pay wage rates, location, job-related skills, experience, and relevant education level or training.
- We are committed to complying with all applicable federal and state minimum wage requirements where applicable.
- If applicable, your recruiter can share more about the specific pay range for your preferred location during the hiring process.
Equal opportunity
- KLA is proud to be an Equal Opportunity Employer.
- Please contact us at talent.acquisition@kla.com or at +1-408-352-2808 to request accommodation.
This listing is sourced directly from Kla's careers page and normalized into a canonical job model.