Bolt Graphics

Bolt Graphics

Site Reliability Engineer

Sunnyvale, CA

No sponsorship$145k-$175kDetected 20 days ago
PythonGoRustC++BashAWSAzureDockerKubernetesLinuxPrometheusSite Reliability EngineeringIncident ResponseAdaptability

About the role

  • This role is mission-critical to maintaining uptime, performance, and operational excellence across compute, storage, and networking environments.
  • Exceptional Linux expertise and advanced automation capabilities are mandatory for success in this role.
  • Continuously monitor developer and production environments and proactively remediate reliability risks.

Responsibilities

  • Design, implement, and operate highly available, fault-tolerant infrastructure and services.
  • Install, maintain, and upgrade server, storage, and networking hardware in office and colocation facilities.
  • Participate in an on-call rotation and lead incident response efforts, including rapid triage, mitigation, and post-incident root cause analysis.
  • Develop, maintain, and continuously improve automation and operational tooling using Bash and Python.
  • Partner closely with engineering teams to support development, testing, and production workloads at scale.

Requirements

  • 5-7 years' experience in managing SRE related functions Expert-level Linux systems administration across complex, production environments (this is a core requirement).
  • Exceptional proficiency in Bash and Python
  • Proven ability to write maintainable automation and diagnostic tooling for large-scale systems.
  • Hands-on experience with virtualization platforms including Proxmox (current), VMware vSphere, and/or OpenShift.
  • Strong experience with containerization technologies (Docker, containerd) and orchestration platforms (Kubernetes).
  • Experience operating workloads in AWS and/or Microsoft Azure environments.
  • Experience implementing observability, monitoring, and alerting using tools such as Prometheus and Grafana.
  • (required): 5-7 years' experience in managing SRE related functions Expert-level Linux systems administration across complex, production environments (this is a core requirement).

Nice to have

  • Familiarity with systems programming languages such as C, C++, Rust, Go, and/or Julia.
  • Relevant certifications such as CompTIA A+, Azure Engineer, or similar are preferred.
  • Active government clearance or the ability to obtain one is required.
  • On-Call & Incident Response Expectations: This role includes participation in an on-call rotation supporting developer and production systems.

Compensation

  • $145,000-$175,000 per year (California).
  • This range represents the anticipated base pay for this role; the final offer may vary based on qualifications, experience, and location.

Company info

  • Bolt Graphics is a semiconductor startup based in Sunnyvale, CA building the fastest and most efficient graphics processors.
  • We pride ourselves on our first principles approach to solving problems.
  • We are energized by our mission to reduce the barrier of entry for content creation and consumption.
  • Our goal is to enable everyone to easily create, simulate and consume immersive experiences as vividly as they can imagine them.
  • Unmute yourself.
  • Test boundaries and get proven right.

Visa & Work Authorization

  • Please note that Bolt Graphics does not currently sponsor candidates for this role

This listing is sourced directly from Bolt Graphics's careers page and normalized into a canonical job model.