Cribl
Senior Site Reliability Engineer
Remote - United States · Senior
Sponsorship not specified$142k-$195kDetected 14 days ago
JavaScriptTypeScriptNode.jsAWSAzureCloud PlatformsTerraformCI/CDLinuxPrometheusGrafanaSite Reliability EngineeringCybersecurityIncident Response
About the role
- We're one of the fastest‑growing private companies and a leading player in a massive, fast‑moving market.
- With a global workforce, we're remote‑first and grounded in a simple idea: software is a people business.
- Cribl is the place where curious, collaborative people can do their best work, grow fast, and bring their full selves to the herd.
Responsibilities
- Diversity drives innovation, enables better decisions to support our customers, and inspires change for the better.
- We're building a culture where differences are valued and welcomed, and we work together to bring out the best in each other.
- Seek out the cause of errors and instability in our production cloud services and drive teams towards better operational excellence
- Help identify and drive down toil with creative innovation and automation
- Proven experience designing, implementing, and operating observability systems for complex cloud-based platforms, with deep knowledge of best practices and a strong drive to implement them leveraging Cribl products.
- Strong knowledge of cloud design patterns for scale, data management, resiliency, etc.
Requirements
- Experience with Configuration Management and Infrastructure as a Code Tools like Terraform (preferred) or Ansible.
- Experience with APM and Observability and related tools such as, New Relic, Splunk, CloudWatch, Prometheus, Grafana/Kibana, Sentry etc.
- Extensive experience with enterprise scale continuous delivery environments.
- Experience with sustainable incident response in a blameless environment.
- Experience with Incident response related tools for instance, PagerDuty, FireHydrant, Blameless etc.
Compensation
- $141,800 - $195,000 USD
Benefits
- Measure and monitor all production systems with an eye towards availability, latency and overall system health
This listing is sourced directly from Cribl's careers page and normalized into a canonical job model.