Site Reliability Engineer III

Posted Yesterday
Be an Early Applicant
Hiring Remotely in Whitehouse, Belfast, Northern Ireland, GBR
Remote or Hybrid
Entry level
Financial Services
The Role
Engineer and operate reliable GCP infrastructure and middleware platforms supporting high-concurrency, ultra-low-latency trading applications. Responsibilities include migrating messaging, service discovery, and data distribution systems; maintaining observability, SLIs, and SLOs; responding to production incidents; reducing toil through automation; supporting disaster recovery and resiliency testing; and mentoring junior engineers.
Summary Generated by Built In

Job Title: Site Reliability Engineer (SRE) III – Platform Engineering & Systems Reliability

The Role: CME Group is seeking a Site Reliability Engineer (SRE) III to engineer reliability for our Google Cloud (GCP) infrastructure, Middleware Platform Engineering team, and core technology foundations powering our Clearing, Risk, and derivatives applications. In this role, you will help build resilient, automated systems that combine ultra-low latency with high-concurrency performance, enabling CME's product teams to innovate safely at scale. You will work alongside senior engineers, mentor junior colleagues, engage in the dynamic operation of production systems, and assist in driving our cloud transformation.

What You Will Do / Key Responsibilities

  • Middleware & Application Architecture: Architect, operate, and support the migration of application platforms—including Messaging (Kafka, RedPanda, MQ, Pub/Sub), Service Discovery (Consul, Vault), and Data Distribution (SFTP/JScape)—to Google Cloud Platform. Manage cluster lifecycles, data replication, RBAC, and workload placement.

  • Observability & Monitoring Fabric: Design, scale, and maintain our observability backbone using tools like OpenTelemetry, Splunk, Prometheus, and Grafana. Establish and continuously improve metrics, logs, alerting strategies, SLIs, and SLOs to enable fast issue detection.

  • Incident Response & Operations: Engage with urgency in live production incidents, take ownership of minor incidents, lead post-mortems, and ensure rapid system recovery.

  • Toil Reduction & Automation: Actively identify operational toil and eliminate manual effort through code, automation, and systematic platform improvements.

  • Resiliency & Testing: Contribute to disaster recovery (DR) strategies, continuous systems resiliency testing, and present reliability improvement suggestions to the Product backlog.

  • Collaboration & Leadership: Lead technical discussions for assigned scope, present solution options, collaborate across functional teams, and mentor junior SRE colleagues.

What We're Looking For

  • Engineering & Scripting Discipline: Programming and scripting skills in high-level languages such as Python, Go, Java, or Bash to construct production-grade tooling.

  • Cloud Native & Systems Fundamentals: Proficiency with Linux-based systems, distributed systems, containerization (Kubernetes/GKE), and public cloud platforms (GCP/GCE).

  • Infrastructure as Code (IaC): Understanding of modern CI/CD patterns and IaC tools such as Terraform, Ansible, or Kubernetes Config Connector (KCC).

  • Networking & Protocols: Knowledge of core systems and networking concepts (TCP/IP, UDP, HTTP, DNS, load balancing, and messaging protocols).

  • AI & Agentic Engineering: Forward-thinking approach to automation, leveraging Generative AI and Agents (e.g., Gemini) to optimize platform operations.

  • Analytical Problem-Solving: Data-driven mindset to troubleshoot complex, non-linear system behaviors in a fast-paced, high-pressure trading ecosystem.

  • Communication & Adaptability: Strategic communication skills to translate technical requirements for cross-functional teams, coupled with an eagerness to learn independently and collaboratively.

Preferred Qualifications / Desirable

  • Observability Stack: Hands-on experience with telemetry tools such as OpenTelemetry, Splunk, Prometheus, and Grafana.

  • Agile Integration: Comfort working within Agile frameworks and collaborative software development lifecycles.

  • Certifications: GCP Professional Cloud Architect, Certified Kubernetes Administrator (CKA), or Certified Kubernetes Application Developer (CKAD).

  • Domain Expertise: Any experience in Financial Markets or other highly regulated, ultra-low latency, high-concurrency environments would be highly beneficial although not essential," 

Why CME Group?

  • Global Significance: Build technology that underpins the integrity of the world's leading derivatives marketplace.

  • Engineering Culture: Flourish in a "code-first" environment that prioritizes systematic, automated solutions over manual intervention.

  • Professional Evolution: Grow your SRE career within an organization actively transforming its approach to production engineering.

  • Competitive Package: Enjoy a robust compensation and benefits structure while working with cutting-edge tech.

Company Benefits:

  • Bonus Programme

  • Equity Programme

  • Employee Stock Purchase Plan (ESPP)

  • Private Medical and Dental coverage

  • Mental Health Benefit Programme

  • Group Pension Plan

  • Income Protection

  • Life Assurance

  • Cycle To Work

  • EV Car Benefit Scheme

  • Gym Membership

  • Family Leave

  • Education Assistance – MBA/Advanced Degree/Bachelor Degree

  • Ongoing Employee Development Training/Certification

  • Hybrid Working

#LI-RK2

#LI-Hybrid

#nijobs.com

CME Group: Where Futures are Made

CME Group is the world’s leading derivatives marketplace. But who we are goes deeper than that. Here, you can impact markets worldwide. Transform industries. And build a career by shaping tomorrow. We invest in your success and you own it – all while working alongside a team of leading experts who inspire you in ways big and small. Problem solvers, difference makers, trailblazers. Those are our people. And we’re looking for more.

At CME Group, we embrace our employees' unique experiences and skills to ensure that everyone’s perspectives are acknowledged and valued. As an equal-opportunity employer, we consider all potential employees without regard to any protected characteristic.

Important Notice: Recruitment fraud is on the rise, with scammers using misleading promises of job offers and interviews to solicit money and personal information from job seekers. CME Group adheres to established procedures designed to maintain trust, confidence and security throughout our recruitment process. Learn more here.

Skills Required

  • Programming or scripting experience with Python, Go, Java, or Bash
  • Experience with Linux-based systems, distributed systems, containerization, Kubernetes or GKE, and public cloud platforms such as GCP or GCE
  • Understanding of CI/CD practices and infrastructure-as-code tools such as Terraform, Ansible, or Kubernetes Config Connector
  • Knowledge of TCP/IP, UDP, HTTP, DNS, load balancing, and messaging protocols
  • Ability to use generative AI and agents, such as Gemini, to improve platform operations
  • Analytical problem-solving skills for troubleshooting complex system behaviors
  • Strategic communication skills and ability to collaborate across functional teams
  • Experience with OpenTelemetry, Splunk, Prometheus, and Grafana
  • Experience working within Agile frameworks and collaborative software development lifecycles
  • GCP Professional Cloud Architect, Certified Kubernetes Administrator, or Certified Kubernetes Application Developer certification
  • Experience in financial markets or other regulated, ultra-low-latency, high-concurrency environments

CME Group Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about CME Group and has not been reviewed or approved by CME Group.

  • Retirement Support U.S. offerings include both a 401(k) and a company-funded cash-balance pension, strengthening long-term financial security. This dual-track structure is highlighted as a notable differentiator among private employers.
  • Leave & Time Off Breadth PTO and holiday schedules are described as generous, with ample time off and carryover commonly highlighted. This breadth of leave meaningfully enhances perceived total rewards.
  • Flexible Benefits A flexible, hybrid work model applies to many roles, increasing day-to-day usability of the package. Flexibility is framed as a standard feature rather than an exception.

CME Group Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Chicago, IL
3,291 Employees

What We Do

As the world's leading derivatives marketplace, CME Group (www.cmegroup.com) is where the world comes to manage risk. CME Group exchanges offer the widest range of global benchmark products across all major asset classes, including futures and options based on interest rates, equity indexes, foreign exchange, energy, agricultural commodities, metals, weather and real estate. CME Group brings buyers and sellers together through its CME Globex® electronic trading platform and its trading facilities in New York and Chicago. CME Group also operates CME Clearing, one of the world’s leading central counterparty clearing provider in the world, which offers clearing and settlement services for exchange-traded contracts, as well as for over-the-counter derivatives transactions through CME ClearPort®. These products and services ensure that businesses everywhere can substantially mitigate counterparty credit risk in both listed and over-the-counter derivatives markets.

Similar Jobs

Affirm Logo Affirm

Staff Software Engineer

Big Data • Fintech • Mobile • Payments • Financial Services
Easy Apply
Remote
UK
2200 Employees
142K-190K Annually

SailPoint Logo SailPoint

Advisory Agentic Technologist

Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Remote or Hybrid
United Kingdom
2461 Employees

Samsara Logo Samsara

Entry Level Tech Sales - Benelux Market (work from the UK, Netherlands, Germany or France - Dutch speaking role)

Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
4 Locations
4000 Employees
In-Office or Remote
2 Locations
2653 Employees

Similar Companies Hiring

Granted Thumbnail
Artificial Intelligence • Healthtech • Insurance • Mobile • Financial Services
New York, New York
23 Employees
Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account