Google

Google

Senior Performance Co-Design Engineer, LLM Serving

Sunnyvale, CA, USA · Senior

Sponsorship not specified$174k-$252kDetected 1 day ago
C++GCPMachine LearningDeep LearningNLPLLMsElectrical EngineeringResearch

> stay_score

odds of building a lasting career here

64Sponsors, lottery-bound
Cap-exempt (no lottery)0
Sponsors this role100
Entry-level history0
PERM / green-card track0
Lottery odds (Level IV)94
Fits your clock70

Sponsors, but it's cap-subject — you still face the weighted lottery (~61% per draw at Level IV). Good if you win; have a cap-exempt backup on your list.

Lottery odds assume a STEM candidate.

Personalize to your clock →

> community_outcomes

No reports yet — be the first to help the next applicant.

About the role

  • - Bachelor's degree in Computer Science, Electrical Engineering, Computer Engineering, a related field, or equivalent practical experience.
  • - 5 years of experience in performance modeling/engineering, computer architecture, co-design, or systems engineering.

Responsibilities

  • Develop and maintain advanced simulation, profiling, and modeling tools to identify bottlenecks, understand key characteristics and project serving workload performance.
  • Partner with model researchers, software and hardware teams to co-design architectural improvements tailored to Large Language Model (LLM) inference latency and throughput.
  • Drive data-backed decisions that influence the roadmap for future TPU/Cloud Silicon architectures.

Requirements

  • Bachelor's degree in Computer Science, Electrical Engineering, Computer Engineering, a related field, or equivalent practical experience.
  • 5 years of experience in performance modeling/engineering, computer architecture, co-design, or systems engineering.
  • Experience programming in C++ or Python.

Nice to have

  • Master's degree or PhD in Electrical Engineering, Computer Engineering or Computer Science, with an emphasis on computer architecture.
  • Experience with hardware/software co-design problems, especially performance analysis and identification at the pre-silicon stage.
  • Experience enabling and optimizing large-scale ML models (e.g., LLMs, large embedding models).
  • Familiarity with accelerator architectures.
  • Google Cloud's mission is to make every business successful through AI by combining cutting-edge technology, infrastructure, and talent.
  • you're shaping the frontier of enterprise and driving the evolution of advanced models.
  • In this role, you will work on analyzing and optimizing the serving performance of emerging models and use cases on our custom hardware.
  • You will also work closely with hardware architects to influence the evolution of Google's custom ML accelerators.

Skills

  • AI/ML software engineers in Cloud bridge the gap between pioneering models and a massive product vehicle reaching billions.

Compensation

  • US: $174000 - $252000 (USD) + 15% bonus target + equity + benefits

Company info

  • We empower Google customers with breakthrough capabilities and insights by delivering AI and Infrastructure at unparalleled scale, efficiency, reliability and velocity.
  • Our customers include Googlers, Google Cloud customers, and billions of Google users worldwide.

This listing is sourced directly from Google's careers page and normalized into a canonical job model.