Relay
Senior Site Reliability Engineer
Vancouver, BC · Senior
Sponsorship not specifiedDetected 166 days ago
TypeScriptNode.jsGitPostgreSQLDynamoDBAWSKubernetesTerraformGitHub ActionsDatadogSite Reliability EngineeringRESTIncident ResponseComplianceLeadershipCommunicationMentoring
About the role
- We do this by replacing financial guesswork with real visibility, transforming cash flow from a constant source of stress into a clear signal owners can use to run stronger, more resilient businesses.
- Your love of making high-impact decisions daily and desire to help shape the future of Relay is going to be crucial.
- Our team is lean, high-impact, and deeply collaborative - we want someone who is always willing to pitch in and isn't afraid to ask for help - You are curious.
Responsibilities
- Join the team building and owning our production infrastructure and CICD pipelines (AWS, Kubernetes, PostgreSQL databases, Terraform, Terragrunt, Github Actions)
- Partner with product teams to balance feature delivery with reliability constraints
- Establish and maintain error budgets for core services
- Collaborate in leading incident response by driving fast mitigation, clear communication, and structured decision-making
- You'd rather build something better than defend something familiar.
- If you require accommodations at any stage of the hiring process, please reach out to your Talent Partner.
- We are looking for someone to help us continue building security into every aspect of our work - and is ready to be on-call for production issues
Requirements
- You have experience as a site reliability engineer working with these technologies: AWS, Kubernetes, Datadog, GitHub, and GitHub Actions
- You have experience owning observability initiatives (logging pipelines, Software Catalog, monitoring strategies, and incident management tooling)
- You have a strong security and operations mindset.
- You are a team player.
- You are curious.
- Experience working in compliant environments such as SOC2 or PCI
Skills
- AWS, Kubernetes, Datadog, GitHub, and GitHub Actions
- You have various levels of experience with Terraform, Terragrunt, Node.js, Typescript
- You have hands-on experience managing and optimizing databases such as Aurora RDS, PostgreSQL, DynamoDB, and ElastiCache
- You are curious. You keep yourself on the bleeding edge of infrastructure best practices
- Bonus Points
- Show us your home lab
- Show us your Github profile!
- Fintech/regulatory experience; Experience working in compliant environments such as SOC2 or PCI
- Experience driving large reliability initiatives across your company
- You've joined a company at its early stages and have seen it through scale
- You have experience working in a fintech startup
- The Interview Process
Compensation
- A take-home case study followed by a 60-minute in-person presentation to members of our SRE team.
- Our Compensation Approach
- We believe Relayers should feel rewarded for the impact they have on our mission and growth.
- Compensation follows impact.
- As impact increases, compensation grows, and we do not limit compensation changes to a once-a-year review cycle.
- The annual salary range for this role is $153,000 CAD to $187,000 CAD.
Benefits
- Build monitoring systems to dynamically assess the infrastructure health
Company info
- We're looking for people who are relentless, curious, and care deeply about the work.
This listing is sourced directly from Relay's careers page and normalized into a canonical job model.