CME Group
Site Reliability Engineer II
New York - 300 Vesey Street, USA · Mid
Sponsorship not specified$94k-$157kDetected 26 days ago
PythonBashGCPCloud PlatformsKubernetesLinuxPrometheusSite Reliability EngineeringMachine LearningLLMsIncident ResponseForecastingCommunicationProblem Solving
About the role
- Role is located in NYC with alternative location Chicago, IL.
- The successful candidate will work alongside senior engineers to learn how we observe, monitor, automate, and improve Production service reliability.
- Join CME Group and play a crucial role in ensuring the stability and performance of our Markets applications while contributing to our GCP migration and AIOps evolution.
Responsibilities
- Work alongside product teams and senior engineers to assist with building out observability, monitoring, and alerting for key services.
- Implement AI-driven reliability solutions, including anomaly detection, predictive alerting, and root cause analysis in production environments.
- Collaborate with engineers and product teams to ensure requirements are understood, planned carefully, and implemented safely.
- Write scripts and tools to reduce toil and improve velocity, including building or integrating intelligent auto-remediation and capacity forecasting systems.
- Support the migration of markets applications to Google Cloud Platform (GCP).
- Collaborate with cross-functional teams to improve system performance and operational efficiency.
Requirements
- AI/ML for Operations: Demonstrated hands-on experience applying AI/ML techniques to improve operational efficiency, reliability, or observability.
- AIOps Platforms: Experience using platforms such as Dynatrace, New Relic, Moogsoft, BigPanda, or integrating open-source tools (e.g., Prometheus with ML models).
- Generative AI Tooling: Experience with LLMs for operations, incident management, or log analysis (e.g., using LangChain, LlamaIndex, or tools like PagerDuty AIOps).
- Traditional Observability: Experience with metrics & monitoring tools like OpenTelemetry, Splunk, Prometheus, and Grafana.
- Systems Architecture: Experience with Kubernetes and knowledge of working with distributed systems.
- Core Concepts: Basic knowledge of networking (HTTP/TCP/UDP/IP) and message-oriented middleware.
- Industry & Process: Experience in financial markets and working in an Agile environment.
Nice to have
- Preferred / Desirable
- Preferred / Desirable Qualifications:
Skills
- Where Futures are Made
- CME Group is the world's leading derivatives marketplace.
- Transform industries.
- And build a career by shaping tomorrow.
- Problem solvers, difference makers, trailblazers.
- Those are our people.
Compensation
- Competitive compensation and benefits package.
Benefits
- Experience with Cloud-based platforms-Google Cloud Platform (GCP), GCE, and/or GKE is a strong bonus.
- Actual salary offered will be dependent on a wide array of factors including but not limited to: relevant experience, skills, education and comparison to internal employees (where relevant).
- Our compensation program also includes an annual target bonus opportunity for all employees, as well as the opportunity to become an owner in the company through our broad-based equity program.
- Through our benefits program, we strive to offer flexibility, value and choice.
- Competitive compensation and benefits package.
Company info
- This is Hybrid role, 2 days on site.
- We are looking for local candidates only.
- Be part of a global leader in financial services technology.
Apply directly at CME Group →Create a free account for alerts like thisView CME Group immigration profile
This listing is sourced directly from CME Group's careers page and normalized into a canonical job model.