CME Group

CME Group

Site Reliability Engineer II

New York - 300 Vesey Street, USA · Mid

Sponsorship not specified$94k-$157kDetected 26 days ago
PythonBashGCPCloud PlatformsKubernetesLinuxPrometheusSite Reliability EngineeringMachine LearningLLMsIncident ResponseForecastingCommunicationProblem Solving

About the role

  • Role is located in NYC with alternative location Chicago, IL.
  • The successful candidate will work alongside senior engineers to learn how we observe, monitor, automate, and improve Production service reliability.
  • Join CME Group and play a crucial role in ensuring the stability and performance of our Markets applications while contributing to our GCP migration and AIOps evolution.

Responsibilities

  • Work alongside product teams and senior engineers to assist with building out observability, monitoring, and alerting for key services.
  • Implement AI-driven reliability solutions, including anomaly detection, predictive alerting, and root cause analysis in production environments.
  • Collaborate with engineers and product teams to ensure requirements are understood, planned carefully, and implemented safely.
  • Write scripts and tools to reduce toil and improve velocity, including building or integrating intelligent auto-remediation and capacity forecasting systems.
  • Support the migration of markets applications to Google Cloud Platform (GCP).
  • Collaborate with cross-functional teams to improve system performance and operational efficiency.

Requirements

  • AI/ML for Operations: Demonstrated hands-on experience applying AI/ML techniques to improve operational efficiency, reliability, or observability.
  • AIOps Platforms: Experience using platforms such as Dynatrace, New Relic, Moogsoft, BigPanda, or integrating open-source tools (e.g., Prometheus with ML models).
  • Generative AI Tooling: Experience with LLMs for operations, incident management, or log analysis (e.g., using LangChain, LlamaIndex, or tools like PagerDuty AIOps).
  • Traditional Observability: Experience with metrics & monitoring tools like OpenTelemetry, Splunk, Prometheus, and Grafana.
  • Systems Architecture: Experience with Kubernetes and knowledge of working with distributed systems.
  • Core Concepts: Basic knowledge of networking (HTTP/TCP/UDP/IP) and message-oriented middleware.
  • Industry & Process: Experience in financial markets and working in an Agile environment.

Nice to have

  • Preferred / Desirable
  • Preferred / Desirable Qualifications:

Skills

  • Where Futures are Made
  • CME Group is the world's leading derivatives marketplace.
  • Transform industries.
  • And build a career by shaping tomorrow.
  • Problem solvers, difference makers, trailblazers.
  • Those are our people.

Compensation

  • Competitive compensation and benefits package.

Benefits

  • Experience with Cloud-based platforms-Google Cloud Platform (GCP), GCE, and/or GKE is a strong bonus.
  • Actual salary offered will be dependent on a wide array of factors including but not limited to: relevant experience, skills, education and comparison to internal employees (where relevant).
  • Our compensation program also includes an annual target bonus opportunity for all employees, as well as the opportunity to become an owner in the company through our broad-based equity program.
  • Through our benefits program, we strive to offer flexibility, value and choice.
  • Competitive compensation and benefits package.

Company info

  • This is Hybrid role, 2 days on site.
  • We are looking for local candidates only.
  • Be part of a global leader in financial services technology.

This listing is sourced directly from CME Group's careers page and normalized into a canonical job model.