Crusoe
Senior Hardware Systems Engineer, Performance
Sunnyvale, CA - US · Senior
Sponsorship not specified$170k-$205kDetected 7 days ago
PythonPlatform EngineeringMachine LearningSystems EngineeringElectrical EngineeringHardware DesignCommunicationProblem Solving
> stay_score
odds of building a lasting career here
37Unrated
Cap-exempt (no lottery)0
Sponsors this role0
Entry-level history0
PERM / green-card track0
Lottery odds (Level IV)94
Fits your clock70
No strong sponsorship signal in the public record yet. In the full product we resolve the exact legal entity and show its filing history with a confidence score — treat as unverified until then.
Lottery odds assume a STEM candidate.
Personalize to your clock →> community_outcomes
No reports yet — be the first to help the next applicant.
About the role
- In this role, you will participate in the full hardware lifecycle - from prototype bring-up to large-scale production while driving automation, deep issue resolution, and reliability across Crusoe Cloud's GPU- and CPU-based infrastructure.
- You will be collaborating with hardware, software, infrastructure, and vendor engineering teams while working across platform bring-up, validation and performance characterization.
Responsibilities
- Drive the end-to-end lifecycle of next-generation compute platforms, including evaluation, bring-up, validation, deployment, and production readiness.
- Build and maintain workload performance profiles and reference configurations that guide how clusters are deployed, tuned, and scaled for specific model families and workload classes.
- Analyze system and workload performance, identify bottlenecks, and work across hardware and software layers to drive improvements.
- Collaborate across hardware, firmware, networking, software, infrastructure, reliability, and operations teams to resolve complex platform issues.
Requirements
- Your work will directly impact Crusoe's ability to deploy and operate sustainable, AI-first compute systems with world-class performance and reliability.
- 5-6+ years of experience in hardware systems engineering, platform engineering, performance engineering, ML systems engineering, infrastructure engineering, or related areas.
- Hands-on experience with large-scale GPU or accelerated computing infrastructure for AI/ML or HPC workloads.
- Hands-on experience with distributed training and/or inference workloads at scale, including parallelism strategies and performance tuning across the hardware/software stack.
- Experience with workload benchmarking, performance profiling, and system performance optimization across hardware and software layers.
- Hands-on experience with system bring-up, validation, performance characterization, and root-cause analysis of complex hardware/software issues.
- Ability to analyze system behavior using telemetry, benchmarks, profiling tools, and other quantitative data.
- Strong analytical and problem-solving skills with the ability to operate effectively in ambiguous and rapidly evolving environments.
- Bachelor's or Master's degree in Electrical Engineering, Computer Engineering, Computer Science, or equivalent experience.
Skills
- Crusoe is on a mission to accelerate the abundance of energy and intelligence.
- The demand for AI compute is boundless, and power is a bottleneck.
Compensation
- Compensation will be paid in the range of $170,000 - $205,000.
- Compensation to be determined by the applicants knowledge, education, and abilities, as well as internal equity and alignment with market data.
- 401(k) with a 100% match up to 4% of salary
Benefits
- Restricted Stock Units in a fast growing, well-funded technology company
- Health insurance package options that include HDHP and PPO, vision, and dental for you and your dependents
- Employer contributions to HSA accounts
- Paid Parental Leave
- Paid life insurance, short-term and long-term disability
- Generous paid time off and holiday schedule
- Cell phone reimbursement
- Tuition reimbursement
Company info
- Strong understanding of modern server and accelerator architectures, including CPU, GPU, memory, storage, networking, and high-speed interconnects such as PCIe, InfiniBand, or NVLink.
- Experience developing automation, testing, diagnostics, or data-analysis frameworks using Python, Shell, or similar languages.
- Experience working across multiple engineering disciplines, including hardware, firmware, software, networking, and infrastructure teams.
- Excellent technical communication skills and experience collaborating with internal engineering teams, customers, and external technology partners.
- Experience influencing hardware or system configuration decisions based on workload performance data (e.g: HW/SW co-design, platform tuning studies).
- Deep experience with RDMA, RoCE, CXL, NVLink or fabric-level performance analysis.
- Experience with inference serving frameworks, training frameworks, or ML compiler/runtime stacks.
- Familiarity with both x86 and ARM-based server platforms.
- Experience building observability, diagnostics, or fleet-level performance and reliability systems.
- Experience introducing new compute technologies into production cloud or large-scale datacenter environments.
- Understanding of infrastructure efficiency, power, cooling, performance-per-dollar, or total cost of ownership considerations.
- Background in sustainable or energy-efficient hardware design practices.
- Advanced certifications or coursework in AI/HPC hardware systems.
Equal opportunity
- Crusoe is an Equal Opportunity Employer.
Visa & Work Authorization
- Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran st
This listing is sourced directly from Crusoe's careers page and normalized into a canonical job model.