Thinkbrg
Site Reliability Engineer
Remote - USA
Work authorization required$130k-$160kDetected 25 days ago
PythonGoRubyFull-Stack DevelopmentGitAWSGCPAzureCloud PlatformsCI/CDGitHub ActionsDatadogSite Reliability EngineeringCybersecurityResearchCommunication
About the role
- The SRE will work closely with software developers and operations teams to improve system reliability, automate processes, and minimize downtime.
- Salary: $130,000 - $160,000 About BRG BRG combines world-leading academic credentials with world-tested business expertise and purpose-built emerging technologies.
- Our unique structure nurtures the interdisciplinary relationships that give us the edge, laying the groundwork for more informed insights and more original, incisive thinking.
Responsibilities
- Design, implement, and maintain scalable and reliable systems in cloud environments such as Azure Cloud Services.
- Provide operational support for full-stack software applications.
- Develop service-level indicators and objectives to automate release validation.
- Manage cloud and database system maintenance, debugging production issues as they arise.
- Partner with security and product teams to define and publish policies, processes, and playbooks to facilitate rapid and effective handling of alerts and incidents.
- Lead incident management processes
- Lead incident management processes; respond to outages and service disruptions promptly.
- Our customers and partners trust us to deliver reliable, first-to-market solutions and safeguard the data we receive.
- We trust our employees, and our culture gives them the freedom to create, collaborate, and grow.
- We are seeking a Site Reliability Engineer to design, build, and maintain highly available systems and infrastructure.
Requirements
- Bachelor's degree in computer science or similar field.
- Proven ability to diagnose and monitor performance and reliability issues across the stack.
- Expertise in Kubernetes.
- Proven experience working with cloud-native infrastructure (Azure Cloud Services, AWS, or GCP).
- Experience working with observability and incident management tools (Datadog, OpsGenie, PagerDuty).
- Experience scripting operating system tasks with Infrastructure as Code.
- Ability to problem-solve in a fast-paced, high-stakes environment.
- Candidate must be able to submit verification of his/her legal right to work in the United States, without company sponsorship.
- BRG combines world-leading academic credentials with world-tested business expertise and purpose-built emerging technologies.
- Five years' experience as a site reliability engineer or similar role.
- Strong programming skills (Golang, Ruby, Python, or similar)
- Relevant industry certifications, such as through the Site Reliability Engineering (SRE) Foundation.
- Impeccable communication skills.
- Salary: $130,000 - $160,000
- About BRG
Compensation
- $130,000 - $160,000
Equal opportunity
- BRG is proud to be an Equal Opportunity Employer.
Visa & Work Authorization
- Candidate must be able to submit verification of his/her legal right to work in the United States, without company sponsorship.
Apply directly at Thinkbrg →Create a free account for alerts like thisView Thinkbrg immigration profile
This listing is sourced directly from Thinkbrg's careers page and normalized into a canonical job model.