MLabs
Senior Site Reliability Engineer (Azure)
United States · Senior · Full-time
No sponsorship$150k-$200kDetected 23 days ago
GoDistributed SystemsAzureCloud PlatformsKubernetesTerraformCI/CDPrometheusSite Reliability EngineeringIncident ResponseSystems EngineeringLeadershipCommunicationCollaborationProblem Solving
About the role
- Scalable Deployments: All customer deployments are verified as repeatable, scalable, and secure.
- Interview Process Recruiter & Technical Screening: Initial HR call followed by an introductory technical interview covering foundational questions.
Responsibilities
- Infrastructure Design: Architect and deploy secure, scalable Azure infrastructure tailored for production-grade distributed systems.
- Automation & IaC: Develop and maintain Terraform-based infrastructure as code to enable repeatable, automated deployments across various environments.
- Platform Enhancement: Build and optimize platform services, APIs, and integrations to extend core system capabilities.
- Cross-Functional Collaboration: Partner with engineering, security, and product teams to deliver enterprise-ready infrastructure solutions.
- Operational Excellence: Drive improvements in reliability, observability, and incident response while providing Tier 2 infrastructure support for customer deployments.
- Develop and maintain Terraform-based infrastructure as code to enable repeatable, automated deployments across various environments.
- Build and optimize platform services, APIs, and integrations to extend core system capabilities.
- Partner with engineering, security, and product teams to deliver enterprise-ready infrastructure solutions.
- Drive improvements in reliability, observability, and incident response while providing Tier 2 infrastructure support for customer deployments.
Requirements
- Ability to transform high-level requirements into scalable, delivered systems.
- Strong technical communication skills with the ability to interface with both engineering teams and non-technical stakeholders.
- Deep knowledge of Azure networking, compute, identity, security, and storage.
- Advanced proficiency with Terraform at production scale.
Nice to have
- Hands-on experience with Kubernetes and container orchestration.
- Familiarity with observability tools such as Prometheus and Grafana.
- Experience with workflow/orchestration platforms like Argo or Spacelift.
- If you do not hear back from us within 4 weeks of your application, please assume that you have not been successful on this occasion.
- We genuinely appreciate your interest and wish you the best in your job search.
- We ensure no discrimination, accessible job adverts, and providing information in accessible formats.
- Our goal is to foster a diverse, inclusive workplace with equal opportunities for all.
- MLabs Ltd collects and processes the personal information you provide such as your contact details, work history, resume, and other relevant data for recruitment purposes only.
Skills
- Azure achieves full feature parity with all other supported cloud environments within the organization's ecosystem.
- A comprehensive evaluation of architectural and technical execution skills.
- Feature Parity: Azure achieves full feature parity with all other supported cloud environments within the organization's ecosystem.
Compensation
- $150K - $200K Our client is seeking a Senior Site Reliability Engineer (Azure) to architect and scale a robust infrastructure foundation for a high-growth distributed systems platform.
- This position is critical for ensuring that the platform operates as a secure, scalable, and production-ready environment capable of supporting complex enterprise use cases and high reliability standards.
- The successful candidate will take a lead role in designing infrastructure from first principles, bridging the gap between product requirements and technical execution.
- This is a high-impact opportunity for a seasoned engineer to build greenfield Azure environments and establish operational excellence across a global ecosystem.
- Infrastructure Design: Architect and deploy secure, scalable Azure infrastructure tailored for production-grade distributed systems.
- Automation & IaC: Develop and maintain Terraform-based infrastructure as code to enable repeatable, automated deployments across various environments.
Benefits
- Comprehensive health insurance and 401k plans (available for US-based employees).
Company info
- Annual incentives based on individual and company milestones.
- At MLabs, we are committed to offer equal opportunities to all candidates.
- If you need any reasonable adjustments during any part of the hiring process or you would like to see the job-advert in an accessible format please let us know at the earliest opportunity by emailing human-resources@mlabs.city.
- Your data may be shared only with clients and trusted partners where necessary for recruitment purposes.
- Due to the high volume of applications we anticipate, we regret that we are unable to provide individual feedback to all candidates.
This listing is sourced directly from MLabs's careers page and normalized into a canonical job model.