Cribl

Cribl

Senior Site Reliability Engineer

Remote - United States · Senior

Sponsorship not specified$142k-$195kDetected 14 days ago
JavaScriptTypeScriptNode.jsAWSAzureCloud PlatformsTerraformCI/CDLinuxPrometheusGrafanaSite Reliability EngineeringCybersecurityIncident Response

About the role

  • We're one of the fastest‑growing private companies and a leading player in a massive, fast‑moving market.
  • With a global workforce, we're remote‑first and grounded in a simple idea: software is a people business.
  • Cribl is the place where curious, collaborative people can do their best work, grow fast, and bring their full selves to the herd.

Responsibilities

  • Diversity drives innovation, enables better decisions to support our customers, and inspires change for the better.
  • We're building a culture where differences are valued and welcomed, and we work together to bring out the best in each other.
  • Seek out the cause of errors and instability in our production cloud services and drive teams towards better operational excellence
  • Help identify and drive down toil with creative innovation and automation
  • Proven experience designing, implementing, and operating observability systems for complex cloud-based platforms, with deep knowledge of best practices and a strong drive to implement them leveraging Cribl products.
  • Strong knowledge of cloud design patterns for scale, data management, resiliency, etc.

Requirements

  • Experience with Configuration Management and Infrastructure as a Code Tools like Terraform (preferred) or Ansible.
  • Experience with APM and Observability and related tools such as, New Relic, Splunk, CloudWatch, Prometheus, Grafana/Kibana, Sentry etc.
  • Extensive experience with enterprise scale continuous delivery environments.
  • Experience with sustainable incident response in a blameless environment.
  • Experience with Incident response related tools for instance, PagerDuty, FireHydrant, Blameless etc.

Compensation

  • $141,800 - $195,000 USD

Benefits

  • Measure and monitor all production systems with an eye towards availability, latency and overall system health

This listing is sourced directly from Cribl's careers page and normalized into a canonical job model.