IonQ

IonQ

Staff Site Reliability Engineer

Santa Clara, California, United States · Staff+ · Contract

Work authorization required$163k-$214kDetected 9 hours ago
AWSSite Reliability EngineeringLLMsCybersecurityIncident ResponseFinancial ModelingLogisticsLeadershipMentoring

About the role

  • We are seeking a Staff Site Reliability Engineer.
  • As Staff SRE Engineer, you set the technical direction for reliability across regions and services.

Responsibilities

  • Production reliability: own service-level objectives, error budgets, and production reliability outcomes end to end, and represent reliability in architecture and scaling decisions.
  • Engineer observability: design and operate the observability stack so production services are fully instrumented and define the standards platform and application teams follow.
  • Govern SLOs and error budgets: define and manage service-level objectives, run regular reviews with service owners, and drive corrective action when services consume error budgets unsafely.
  • Drive resilience: design and execute chaos experiments and validate that failure modes are covered by tested safeguards.
  • Disaster recovery: own disaster-recovery testing and failover validation against defined recovery objectives and turn exercise findings into architectural and operational improvements.
  • Cloud security posture: co-own cloud security posture management, runtime vulnerability detection, and configuration-compliance monitoring with DevSecOps.
  • Data, streaming, and AI Ops: own reliability of stateful and streaming services, capacity planning and rightsizing, and autonomous agents for triage, predictive alerting, remediation, and self-healing.
  • own service-level objectives, error budgets, and production reliability outcomes end to end, and represent reliability in architecture and scaling decisions.
  • design and operate the observability stack so production services are fully instrumented and define the standards platform and application teams follow.
  • define and manage service-level objectives, run regular reviews with service owners, and drive corrective action when services consume error budgets unsafely.

Requirements

  • 7+ years of production engineering experience with recent hands-on reliability work.
  • Hands-on, recent experience operating large-scale, fault-tolerant production systems on AWS or GCP.
  • Due to applicable export control laws and regulations, candidates must be a U.S. citizen or national, U.S. permanent resident (i.e., current Green Card holder), or lawfully admitted into the U.S. as a refugee or granted asylum.

Nice to have

  • Proven production experience with cloud security posture management, runtime vulnerability detection, and workload protection across cloud and distributed environments.
  • Strong experience prioritizing risk using identity, workload, and exposure-path context to focus remediation on issues that materially increase attack likelihood and operational impact.
  • Experience with autonomous remediation and self-healing workflows powered by AIOps, including Amazon Bedrock Agent Core or equivalent agentic automation frameworks.
  • Hands-on experience in capacity management, resource rightsizing, efficiency engineering, and practical cost optimization based on FinOps principles.
  • We empower employees to thrive by fostering a culture of autonomy, productivity, and respect.
  • US Technical Jobs.
  • The position you are applying for will require access to technology that is subject to U.S. export control and government contract restrictions.
  • US Non-Technical Jobs.

Skills

  • In 2025, the company achieved 99.99% two-qubit gate fidelity, setting a world record in quantum computing performance.
  • IonQ is making quantum platforms more accessible and impactful than ever before.

Compensation

  • Posted base salary figures are subject to change as new market data becomes available.

Benefits

  • Lead incident response: define the incident process and serve as incident commander for the highest-severity incidents, including security incidents within the coverage window.
  • Run on-call and escalation: establish and manage rotations and escalation paths that provide continuous coverage with clean follow-the-sun handoffs.
  • will vary based on individual factors such as education, qualifications, and experience of the final candidate(s), specific
  • Our benefits include comprehensive medical, dental, and vision plans, matching 401(k), unlimited PTO and paid holidays, parental/adoption leave, legal insurance, and a home technology stipend.
  • Details of participation in these benefit plans will be provided when a candidate receives an offer of employment.
  • We are committed to equity and justice.
  • office location, and calibration against relevant market data and internal team equity.

Company info

  • mentor engineers at different seniority levels, set standards adopted across teams, and align Architecture, DevSecOps, Cloud Operations, and Product Development behind a shared reliability roadmap.

Visa & Work Authorization

  • export control and government contract restrictions

This listing is sourced directly from IonQ's careers page and normalized into a canonical job model.