Kla

Kla

High Performance Compute (HPC) Software Engineer – HPC SW Systems

Ann Arbor, MI

Sponsorship not specified$106k-$180kDetected 30 days ago
JavaC++Node.jsAlgorithmsDockerKubernetesLinuxElectrical EngineeringAR/VRResearchCommunicationCollaboration

About the role

  • Company Overview KLA is a global leader in diversified electronics for the semiconductor manufacturing ecosystem.
  • No laptop, smartphone, wearable device, voice-controlled gadget, flexible screen, VR device or smart car would have made it into your hands without us.

Responsibilities

  • Design, develop, and optimize HPC software running on large-scale Linux clusters, including distributed and parallel workloads (MPI, multithreading, GPU-accelerated pipelines, containerized workloads).
  • Optimize application performance and power utilization across CPU, memory, storage, and network subsystem, with attention to throughput, latency, and scaling behavior.
  • Collaborate with hardware and systems teams to define HPC node, storage, and interconnect requirements based on software and algorithm needs.
  • Contribute to best practices, and design reviews for new platforms and refresh cycles.

Requirements

  • Bachelor's or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent practical experience.
  • Proficiency in Java and/or C++ and/or other system-level or performance-oriented languages.
  • Doctorate (Academic) Degree and 0 years related work experience
  • Master's Level Degree and related work experience of 3 years
  • Bachelor's Level Degree and related work experience of 5 years

Nice to have

  • Experience with containerized HPC environments (Docker, Singularity/Apptainer, Kubernetes in HPC contexts).
  • Familiarity with high-speed interconnects, storage architectures, and performance benchmarking.
  • Exposure to rack integration, including cabling, power distribution, cooling, and system bring-up.
  • Experience in semiconductor, manufacturing, or high-reliability systems environments.
  • Ability to reason about system reliability, MTBF/MTBA, and failure modes in large compute installations.
  • What Makes This Role Unique at KLA
  • Work on mission-critical HPC platforms that directly impact semiconductor manufacturing capability.
  • See your work deployed at scale in real production tools-not just in the data center.

Skills

  • CPUs, memory hierarchies, storage, networking (Ethernet / InfiniBand).
  • Practical experience working with clusters, servers, or rack-scale systems in lab or production environments.
  • Strong debugging skills across software, OS, and hardware boundaries.
  • Virtually every electronic device in the world is produced using our technologies.

Compensation

  • $105,900.00 - $180,000.00 Annually
  • Our pay ranges are determined by role, level, and location.
  • The range displayed reflects the pay for this position in the primary location identified in this posting.
  • Actual pay depends on several factors, including state minimum pay wage rates, location, job-related skills, experience, and relevant education level or training.
  • We are committed to complying with all applicable federal and state minimum wage requirements where applicable.
  • If applicable, your recruiter can share more about the specific pay range for your preferred location during the hiring process.

Equal opportunity

  • KLA is proud to be an Equal Opportunity Employer.
  • Please contact us at talent.acquisition@kla.com or at +1-408-352-2808 to request accommodation.

This listing is sourced directly from Kla's careers page and normalized into a canonical job model.