Cisco
Senior Site Reliability Engineer, Production Engineer - ThousandEyes
San Francisco, California, US · Senior · Full-time
Sponsorship not specified$165k-$241kDetected 14 hours ago
PythonGoDistributed SystemsAWSKubernetesLinuxPrometheusSite Reliability EngineeringIncident ResponseCiscoCommunicationCollaborationArgo CD
> stay_score
odds of building a lasting career here
37Unrated
Cap-exempt (no lottery)0
Sponsors this role0
Entry-level history0
PERM / green-card track0
Lottery odds (Level IV)94
Fits your clock70
No strong sponsorship signal in the public record yet. In the full product we resolve the exact legal entity and show its filing history with a confidence score — treat as unverified until then.
Lottery odds assume a STEM candidate.
Personalize to your clock →> community_outcomes
No reports yet — be the first to help the next applicant.
About the role
- The application window is expected to close on: 09/01/2026 Job posting may be removed earlier if the position is filled or if a sufficient number of applications are received.
- This role follows a hybrid work model, with in-office attendance expected once a week in San Francisco, Seattle, Austin or New York.
- Leveraging AI and an unparalleled set of cloud, internet, and enterprise network telemetry data, ThousandEyes enables IT teams to proactively detect, diagnose, and resolve issues before they impact end-user experiences.
Responsibilities
- Collaborate with software engineers to optimize architecture and services for availability, latency, performance, and reliability using cloud-native tools.
- Design and implement scalable operations tooling to support platform growth and scaling across multiple regions.
- Design, deploy, and maintain AWS cloud-native services that are elastic and resilient to failure.
- Develop automation solutions for scalable service and platform operations, including deployment, scale testing, graceful failure, and chaos testing.
- Manage a rapidly growing infrastructure capable of handling substantial daily data volumes, emphasizing operations/infrastructure/everything as code.
- You will design and manage large-scale, highly available distributed systems in the cloud, collaborating directly with application development teams to enhance the reliability, performance, and security of our platform.
Requirements
- 5+ years of experience in a related role
- Proficiency in software development with languages such as Python or Go
- Strong understanding of Unix/Linux systems, including kernel, system libraries, file systems, and client-server protocols
- Knowledge of Site Reliability principles: Incident Response, Change Management, Distributed Systems, Deployment Strategies, and SLOs
Nice to have
- Familiarity with procedures for operating a large-scale, highly available enterprise platform
- Excellent communication and documentation skills
- Expert-level knowledge of Kubernetes and its ecosystem
- In-depth knowledge of cloud providers, preferably AWS
- Additional paid time away may be requested to deal with critical or emergency issues for family members
- Optional 10 paid days per full calendar year to volunteer
- 1.5% of incentive target for each 1% of attainment between 50% and 75%;
- 1% of incentive target for each 1% of attainment between 75% and 100%
Skills
- Meet the Team
Compensation
- The starting salary range posted for this position is $165,000.00 to $241,400.00 and reflects the projected salary range for new hires in this position in U.S. and/or Canada locations, not including incentive compensation*, equity, or benef
Company info
- Your Impact We are seeking a skilled Senior Site Reliability Engineer (SRE) in Production Engineering with a strong background in SaaS and operations.
- At Cisco, we're revolutionizing how data and infrastructure connect and protect organizations in the AI era - and beyond.
- These solutions provide customers with unparalleled security, visibility, and insights across the entire digital footprint.
- We are seeking a skilled Senior Site Reliability Engineer (SRE) in Production Engineering with a strong background in SaaS and operations.
This listing is sourced directly from Cisco's careers page and normalized into a canonical job model.