Bolt Graphics
Site Reliability Engineer
Sunnyvale, CA
No sponsorship$145k-$175kDetected 20 days ago
PythonGoRustC++BashAWSAzureDockerKubernetesLinuxPrometheusSite Reliability EngineeringIncident ResponseAdaptability
About the role
- This role is mission-critical to maintaining uptime, performance, and operational excellence across compute, storage, and networking environments.
- Exceptional Linux expertise and advanced automation capabilities are mandatory for success in this role.
- Continuously monitor developer and production environments and proactively remediate reliability risks.
Responsibilities
- Design, implement, and operate highly available, fault-tolerant infrastructure and services.
- Install, maintain, and upgrade server, storage, and networking hardware in office and colocation facilities.
- Participate in an on-call rotation and lead incident response efforts, including rapid triage, mitigation, and post-incident root cause analysis.
- Develop, maintain, and continuously improve automation and operational tooling using Bash and Python.
- Partner closely with engineering teams to support development, testing, and production workloads at scale.
Requirements
- 5-7 years' experience in managing SRE related functions Expert-level Linux systems administration across complex, production environments (this is a core requirement).
- Exceptional proficiency in Bash and Python
- Proven ability to write maintainable automation and diagnostic tooling for large-scale systems.
- Hands-on experience with virtualization platforms including Proxmox (current), VMware vSphere, and/or OpenShift.
- Strong experience with containerization technologies (Docker, containerd) and orchestration platforms (Kubernetes).
- Experience operating workloads in AWS and/or Microsoft Azure environments.
- Experience implementing observability, monitoring, and alerting using tools such as Prometheus and Grafana.
- (required): 5-7 years' experience in managing SRE related functions Expert-level Linux systems administration across complex, production environments (this is a core requirement).
Nice to have
- Familiarity with systems programming languages such as C, C++, Rust, Go, and/or Julia.
- Relevant certifications such as CompTIA A+, Azure Engineer, or similar are preferred.
- Active government clearance or the ability to obtain one is required.
- On-Call & Incident Response Expectations: This role includes participation in an on-call rotation supporting developer and production systems.
Compensation
- $145,000-$175,000 per year (California).
- This range represents the anticipated base pay for this role; the final offer may vary based on qualifications, experience, and location.
Company info
- Bolt Graphics is a semiconductor startup based in Sunnyvale, CA building the fastest and most efficient graphics processors.
- We pride ourselves on our first principles approach to solving problems.
- We are energized by our mission to reduce the barrier of entry for content creation and consumption.
- Our goal is to enable everyone to easily create, simulate and consume immersive experiences as vividly as they can imagine them.
- Unmute yourself.
- Test boundaries and get proven right.
Visa & Work Authorization
- Please note that Bolt Graphics does not currently sponsor candidates for this role
Apply directly at Bolt Graphics →Create a free account for alerts like thisView Bolt Graphics immigration profile
This listing is sourced directly from Bolt Graphics's careers page and normalized into a canonical job model.