Senior Performance Co-Design Engineer, LLM Serving
Sunnyvale, CA, USA · Senior
Sponsorship not specified$174k-$252kDetected 1 day ago
C++GCPMachine LearningDeep LearningNLPLLMsElectrical EngineeringResearch
> stay_score
odds of building a lasting career here
64Sponsors, lottery-bound
Cap-exempt (no lottery)0
Sponsors this role100
Entry-level history0
PERM / green-card track0
Lottery odds (Level IV)94
Fits your clock70
Sponsors, but it's cap-subject — you still face the weighted lottery (~61% per draw at Level IV). Good if you win; have a cap-exempt backup on your list.
Lottery odds assume a STEM candidate.
Personalize to your clock →> community_outcomes
No reports yet — be the first to help the next applicant.
About the role
- - Bachelor's degree in Computer Science, Electrical Engineering, Computer Engineering, a related field, or equivalent practical experience.
- - 5 years of experience in performance modeling/engineering, computer architecture, co-design, or systems engineering.
Responsibilities
- Develop and maintain advanced simulation, profiling, and modeling tools to identify bottlenecks, understand key characteristics and project serving workload performance.
- Partner with model researchers, software and hardware teams to co-design architectural improvements tailored to Large Language Model (LLM) inference latency and throughput.
- Drive data-backed decisions that influence the roadmap for future TPU/Cloud Silicon architectures.
Requirements
- Bachelor's degree in Computer Science, Electrical Engineering, Computer Engineering, a related field, or equivalent practical experience.
- 5 years of experience in performance modeling/engineering, computer architecture, co-design, or systems engineering.
- Experience programming in C++ or Python.
Nice to have
- Master's degree or PhD in Electrical Engineering, Computer Engineering or Computer Science, with an emphasis on computer architecture.
- Experience with hardware/software co-design problems, especially performance analysis and identification at the pre-silicon stage.
- Experience enabling and optimizing large-scale ML models (e.g., LLMs, large embedding models).
- Familiarity with accelerator architectures.
- Google Cloud's mission is to make every business successful through AI by combining cutting-edge technology, infrastructure, and talent.
- you're shaping the frontier of enterprise and driving the evolution of advanced models.
- In this role, you will work on analyzing and optimizing the serving performance of emerging models and use cases on our custom hardware.
- You will also work closely with hardware architects to influence the evolution of Google's custom ML accelerators.
Skills
- AI/ML software engineers in Cloud bridge the gap between pioneering models and a massive product vehicle reaching billions.
Compensation
- US: $174000 - $252000 (USD) + 15% bonus target + equity + benefits
Company info
- We empower Google customers with breakthrough capabilities and insights by delivering AI and Infrastructure at unparalleled scale, efficiency, reliability and velocity.
- Our customers include Googlers, Google Cloud customers, and billions of Google users worldwide.
This listing is sourced directly from Google's careers page and normalized into a canonical job model.