Claryo
Integration Reliability Engineer
San Francisco
Sponsorship not specifiedDetected 71 days ago
Distributed SystemsAWSGCPAzureCloud PlatformsKubernetesLinuxPrometheusGrafanaSite Reliability EngineeringKafkaData EngineeringIncident ResponseComplianceVoIP
About the role
- This role is responsible for making systems observable, diagnosable, and repeatable as we scale across deployments.
- We welcome teammates of all backgrounds and don't discriminate based on race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.
Requirements
- 3+ years of experience in SRE, infrastructure, or distributed systems
- Experience operating systems in production environments
- Experience with:
- Ability to debug issues across multiple layers of the stack (infra → services → network)
Nice to have
- Experience with multi-site or edge deployments
- Experience with event-driven systems (Kafka or similar)
- Familiarity with video or streaming systems (RTSP, WebRTC)
- Experience working with hardware-integrated systems
- Exposure to security/compliance frameworks (SOC2, ISO27001, etc.)
- US citizen/ permanent resident
- Located in SFBAY or NY area
- We're scaling from a small number of deployments to many, and this role is critical to making the following happen:
Skills
- Build and maintain monitoring, alerting, and observability systems
- Define and improve incident response, severity levels, and on-call processes
- Improve deployment and bring-up workflows across facilities
- Diagnose issues across infrastructure, networking, and distributed systems
- Improve system visibility, debugging, and operational tooling
- Help make deployments repeatable and scalable across sites
Equal opportunity
- We're an equal opportunity employer that values diversity and inclusion.
- Repeatable deployments across real-world facilities
Visa & Work Authorization
- US citizen/ permanent resident
This listing is sourced directly from Claryo's careers page and normalized into a canonical job model.