Crusoe
Staff Hardware Systems Engineer
San Francisco, CA - US · Staff+
Sponsorship not specified$208k-$253kDetected 12 days ago
PythonMachine LearningSystems EngineeringElectrical EngineeringHardware DesignCommunicationCollaborationProblem Solving
About the role
- In this role, you will take ownership of the full hardware lifecycle-from prototype bring-up to large-scale production-while driving automation, deep issue resolution, and reliability across Crusoe Cloud's GPU- and CPU-based infrastructure.
- PCIe (link training, topology, performance issues)
Responsibilities
- You will work closely with cross-functional teams to support, debug, and improve hardware platforms at scale, with a particular focus on PCIe, InfiniBand, and NVMe/storage, which have been identified as essential areas for deeper expertise.
- Drive the full hardware development and sustaining lifecycle, including feasibility, bring-up, validation, deployment, and ongoing production support.
- Collaborate with mechanical, thermal, firmware, software, and manufacturing teams to resolve system-level issues and enable stable production operation.
- Technical background in digital and analog design, server architecture, and high-performance compute hardware.
Requirements
- Your work will directly impact Crusoe's ability to deploy and operate sustainable, AI-first compute systems with world-class performance and reliability.
- 8-10+ years of experience in hardware development, validation, sustaining engineering, or production engineering.
- Deep proficiency in hardware bring-up, board-level debugging, and system-level validation.
- Bachelor's or Master's degree in Electrical Engineering, Computer Engineering, or equivalent experience.
Nice to have
- Experience designing or optimizing GPU-to-GPU communication architectures for AI/ML workloads.
- Direct experience integrating NVLink or other next-generation GPU interconnect technologies.
- Familiarity with cutting-edge GPU architectures and how to leverage them in AI/HPC environments.
- Expertise supporting or designing systems across both ARM and x86 server architectures.
- Advanced certifications or coursework in AI/HPC hardware systems.
Skills
- Crusoe is on a mission to accelerate the abundance of energy and intelligence.
- The demand for AI compute is boundless, and power is a bottleneck.
Compensation
- Compensation will be paid in the range of $208,000 - $253,000 + Bonus.
- Compensation to be determined by the applicant's education, experience, knowledge, skills, and abilities, as well as internal equity and alignment with market data.
Benefits
- Restricted Stock Units
- Paid time off & paid holidays
- Comprehensive health, dental & vision insurance
- Employer contributions to HSA account
- Paid parental leave
- Paid life insurance, short-term and long-term disability
- Professional development & tuition reimbursement
- Mental health & wellness support
- Commuter benefits (parking & transit)
- Cell phone stipend
- 401(k) Retirement plan with company match up to 4% of salary
Company info
- Strong hands-on expertise in PCIe, InfiniBand, and NVMe/storage debugging and development.
- Ability to design and implement automation frameworks for hardware testing (Python, Shell, or similar).
- Experience working across thermal, mechanical, firmware, and software functions in multidisciplinary environments.
- Strong analytical and problem-solving skills with a data-driven approach.
- Excellent communication and collaboration skills for working with internal teams and external partners.
Equal opportunity
- Crusoe is an Equal Opportunity Employer.
Visa & Work Authorization
- Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran st
This listing is sourced directly from Crusoe's careers page and normalized into a canonical job model.