Thinkbrg

Thinkbrg

Site Reliability Engineer

Remote - USA

Work authorization required$130k-$160kDetected 25 days ago
PythonGoRubyFull-Stack DevelopmentGitAWSGCPAzureCloud PlatformsCI/CDGitHub ActionsDatadogSite Reliability EngineeringCybersecurityResearchCommunication

About the role

  • The SRE will work closely with software developers and operations teams to improve system reliability, automate processes, and minimize downtime.
  • Salary: $130,000 - $160,000 About BRG BRG combines world-leading academic credentials with world-tested business expertise and purpose-built emerging technologies.
  • Our unique structure nurtures the interdisciplinary relationships that give us the edge, laying the groundwork for more informed insights and more original, incisive thinking.

Responsibilities

  • Design, implement, and maintain scalable and reliable systems in cloud environments such as Azure Cloud Services.
  • Provide operational support for full-stack software applications.
  • Develop service-level indicators and objectives to automate release validation.
  • Manage cloud and database system maintenance, debugging production issues as they arise.
  • Partner with security and product teams to define and publish policies, processes, and playbooks to facilitate rapid and effective handling of alerts and incidents.
  • Lead incident management processes
  • Lead incident management processes; respond to outages and service disruptions promptly.
  • Our customers and partners trust us to deliver reliable, first-to-market solutions and safeguard the data we receive.
  • We trust our employees, and our culture gives them the freedom to create, collaborate, and grow.
  • We are seeking a Site Reliability Engineer to design, build, and maintain highly available systems and infrastructure.

Requirements

  • Bachelor's degree in computer science or similar field.
  • Proven ability to diagnose and monitor performance and reliability issues across the stack.
  • Expertise in Kubernetes.
  • Proven experience working with cloud-native infrastructure (Azure Cloud Services, AWS, or GCP).
  • Experience working with observability and incident management tools (Datadog, OpsGenie, PagerDuty).
  • Experience scripting operating system tasks with Infrastructure as Code.
  • Ability to problem-solve in a fast-paced, high-stakes environment.
  • Candidate must be able to submit verification of his/her legal right to work in the United States, without company sponsorship.
  • BRG combines world-leading academic credentials with world-tested business expertise and purpose-built emerging technologies.
  • Five years' experience as a site reliability engineer or similar role.
  • Strong programming skills (Golang, Ruby, Python, or similar)
  • Relevant industry certifications, such as through the Site Reliability Engineering (SRE) Foundation.
  • Impeccable communication skills.
  • Salary: $130,000 - $160,000
  • About BRG

Compensation

  • $130,000 - $160,000

Equal opportunity

  • BRG is proud to be an Equal Opportunity Employer.

Visa & Work Authorization

  • Candidate must be able to submit verification of his/her legal right to work in the United States, without company sponsorship.

This listing is sourced directly from Thinkbrg's careers page and normalized into a canonical job model.