MX
Sr. Site Reliability Engineer
Lehi, Utah, United States · Senior
Sponsorship not specifiedDetected 10 hours ago
PythonJavaGoRubyBashRailsPostgreSQLRedisAWSKubernetesTerraformPrometheusGrafanaDatadogDevOpsSite Reliability EngineeringRESTRabbitMQIncident ResponseEmbedded SystemsLeadership
> stay_score
odds of building a lasting career here
16Unrated
Cap-exempt (no lottery)0
Sponsors this role0
Entry-level history0
PERM / green-card track0
Lottery odds40
Fits your clock70
No strong sponsorship signal in the public record yet. In the full product we resolve the exact legal entity and show its filing history with a confidence score — treat as unverified until then.
Lottery odds assume a STEM candidate.
Personalize to your clock →> community_outcomes
No reports yet — be the first to help the next applicant.
About the role
- Like many startups, we've navigated real growth challenges - and we've come out stronger on the other side.
- Today, MX is in a phase of renewed momentum and scale, with a solid foundation and a clear vision for what's next.
- This is a place where thoughtful execution matters, innovation is encouraged, and individuals have real ownership over their work.
Responsibilities
- Build and operate an observability control plane: automate baseline monitors, dashboards, and tagging standards through the Datadog API and Terraform.
- Validate, don't own. Service owners keep their alerts and dashboards
- Build self-serve onboarding so new services get baseline observability on day one, without a multi-week embed.
- We build technology that helps banks, credit unions, and fintechs deliver smarter, more intuitive financial experiences to millions of people.
- We give people the space to question assumptions, design better solutions, and help shape how the company grows.
- We're building a new observability function that runs the way we run incident response: the system does the heavy lifting, and people handle judgment, customers, and the exceptions.
- As a Senior Observability Engineer, you build and operate an observability control plane.
- you raise the bar for every team through standards and automation instead of building each team's dashboards by hand.
- Validate, don't own.
Requirements
- Fintech experience with MX-like architectures
Nice to have
- 5+ years automation-first engineering in Python, Bash, Go, and/or Terraform, plus Kubernetes proficiency
- Preferred Requirements
- Datadog preferred
- strong Grafana/Prometheus, Splunk, or New Relic experience counts if you can ramp on Datadog fast
Benefits
- You scaffold baselines, score coverage, and turn every real incident into the detection the platform should have caught.
- Escalate to engineering managers when coverage fails or an owner is missing.
- Own the monthly observability and service-catalog health report: departed owners, stale dashboards, services with no monitors, SLO gaps, and coverage trends.
- Governance and reporting: you can produce a monthly health and compliance report leadership reads (orphans, stale entries, gaps, trends)
Company info
- Our culture values curiosity, accountability, and impact.
- Our infrastructure powers financial applications used by millions of people and processes billions of transactions for major financial institutions, and customers feel every second of downtime.
- At MX, we are a high-performance organization that thrives on trust and results.
This listing is sourced directly from MX's careers page and normalized into a canonical job model.