The Trade Desk
Senior Software Engineer - Observability & IRM
Boulder; Denver; Seattle · Senior · Full-time
Stay score
odds of building a lasting career here
Sponsors, but it's cap-subject — you still face the weighted lottery (~61% per draw at Level IV). Good if you win; have a cap-exempt backup on your list.
Lottery odds assume a STEM candidate.
Personalize to your clock →Employer immigration record
from this employer's Department of Labor filings
Green-card filing pattern in this occupation
Green-card intent detected
Green-card follow-through: 50%
Files H-1B transfers
Sourced from Department of Labor LCA, PERM and prevailing-wage disclosure data. Employer matching is by name, so figures may be split across an employer's legal entities. Absence of a filing means none appears in our copy of the data, not that none exists.
Community outcomes
No reports yet — be the first to help the next applicant.
About the role
- The Incident Response Services (IRS) taskforce focuses on the on-call experience.
- Help evaluate and migrate our logging stack
- Participate in the re-evaluation of our logging vendor and collection architecture
Responsibilities
- By making it more transparent, effective, and responsible, we help support trusted journalism, quality entertainment, and creators worldwide.
- The scale of our platform brings unique technical challenges - from processing massive datasets in real time to building systems that operate reliably on a global scale.
- The team is responsible for making incidents easier to detect, manage, and optimize using historical data points information.
- Build and maintain automation around the incident lifecycle: alerting, escalation, incident channels, retros, and SLA tracking
- Alert quality tooling - Build the systems that give engineers better signal and less noise - smarter routing, better grouping, tighter feedback loops between alerts and the teams that own them
Requirements
- Comfort working across the stack - this role touches distributed systems, Kubernetes, observability pipelines, and web-based tooling
- Familiarity with observability concepts: logging, alerting, on-call workflows
- What we care about is that you can learn quickly and find solutions to complex problems using the optimum tools for the job.
- What you know is less important than how well you learn and innovate.
Skills
- Experience with Grafana, Prometheus, or similar observability tools
- Familiarity with Sumo Logic or other log management platforms
- Prior work on developer portals or service catalog tooling (Backstage, OpsLevel, etc.)
- Experience with Kubernetes at scale
- A deep understanding of HunnyPt
- The Trade Desk does not accept unsolicited resumes from search firm recruiters.
- The Trade Desk is an equal opportunity employer.
- All aspects of employment will be based on merit, competence, performance, and business needs.
Compensation
- In accordance with various US state laws, the range provided is the Trade Desk's reasonable estimate of the base compensation for this role.
Company info
- The Service Excellence (SE) team owns the tools and infrastructure that help engineers at The Trade Desk understand and operate production systems.
- What you will work on:
- Incident management tooling
- Backstage/Service catalog - Extend our internal developer portal with K8s integrations, maturity models, and SLO adoption tooling
- Who you are:
- Experience building and operating production infrastructure or internal developer tooling
This listing is sourced directly from The Trade Desk's careers page and normalized into a canonical job model.