Relay

Relay

Senior Site Reliability Engineer

Vancouver, BC · Senior

Sponsorship not specifiedDetected 166 days ago
TypeScriptNode.jsGitPostgreSQLDynamoDBAWSKubernetesTerraformGitHub ActionsDatadogSite Reliability EngineeringRESTIncident ResponseComplianceLeadershipCommunicationMentoring

About the role

  • We do this by replacing financial guesswork with real visibility, transforming cash flow from a constant source of stress into a clear signal owners can use to run stronger, more resilient businesses.
  • Your love of making high-impact decisions daily and desire to help shape the future of Relay is going to be crucial.
  • Our team is lean, high-impact, and deeply collaborative - we want someone who is always willing to pitch in and isn't afraid to ask for help - You are curious.

Responsibilities

  • Join the team building and owning our production infrastructure and CICD pipelines (AWS, Kubernetes, PostgreSQL databases, Terraform, Terragrunt, Github Actions)
  • Partner with product teams to balance feature delivery with reliability constraints
  • Establish and maintain error budgets for core services
  • Collaborate in leading incident response by driving fast mitigation, clear communication, and structured decision-making
  • You'd rather build something better than defend something familiar.
  • If you require accommodations at any stage of the hiring process, please reach out to your Talent Partner.
  • We are looking for someone to help us continue building security into every aspect of our work - and is ready to be on-call for production issues

Requirements

  • You have experience as a site reliability engineer working with these technologies: AWS, Kubernetes, Datadog, GitHub, and GitHub Actions
  • You have experience owning observability initiatives (logging pipelines, Software Catalog, monitoring strategies, and incident management tooling)
  • You have a strong security and operations mindset.
  • You are a team player.
  • You are curious.
  • Experience working in compliant environments such as SOC2 or PCI

Skills

  • AWS, Kubernetes, Datadog, GitHub, and GitHub Actions
  • You have various levels of experience with Terraform, Terragrunt, Node.js, Typescript
  • You have hands-on experience managing and optimizing databases such as Aurora RDS, PostgreSQL, DynamoDB, and ElastiCache
  • You are curious. You keep yourself on the bleeding edge of infrastructure best practices
  • Bonus Points
  • Show us your home lab
  • Show us your Github profile!
  • Fintech/regulatory experience; Experience working in compliant environments such as SOC2 or PCI
  • Experience driving large reliability initiatives across your company
  • You've joined a company at its early stages and have seen it through scale
  • You have experience working in a fintech startup
  • The Interview Process

Compensation

  • A take-home case study followed by a 60-minute in-person presentation to members of our SRE team.
  • Our Compensation Approach
  • We believe Relayers should feel rewarded for the impact they have on our mission and growth.
  • Compensation follows impact.
  • As impact increases, compensation grows, and we do not limit compensation changes to a once-a-year review cycle.
  • The annual salary range for this role is $153,000 CAD to $187,000 CAD.

Benefits

  • Build monitoring systems to dynamically assess the infrastructure health

Company info

  • We're looking for people who are relentless, curious, and care deeply about the work.

This listing is sourced directly from Relay's careers page and normalized into a canonical job model.