CME Group Logo

CME Group

Site Reliability Engineer III

Posted 4 Hours Ago
Be an Early Applicant
Remote or Hybrid
Hiring Remotely in Whitehouse, Belfast, Northern Ireland, GBR
Entry level
Remote or Hybrid
Hiring Remotely in Whitehouse, Belfast, Northern Ireland, GBR
Entry level
Engineer and operate reliable GCP infrastructure and middleware platforms supporting high-concurrency, ultra-low-latency trading applications. Responsibilities include migrating messaging, service discovery, and data distribution systems; maintaining observability, SLIs, and SLOs; responding to production incidents; reducing toil through automation; supporting disaster recovery and resiliency testing; and mentoring junior engineers.
The summary above was generated by AI

Job Title: Site Reliability Engineer (SRE) III – Platform Engineering & Systems Reliability

The Role: CME Group is seeking a Site Reliability Engineer (SRE) III to engineer reliability for our Google Cloud (GCP) infrastructure, Middleware Platform Engineering team, and core technology foundations powering our Clearing, Risk, and derivatives applications. In this role, you will help build resilient, automated systems that combine ultra-low latency with high-concurrency performance, enabling CME's product teams to innovate safely at scale. You will work alongside senior engineers, mentor junior colleagues, engage in the dynamic operation of production systems, and assist in driving our cloud transformation.

What You Will Do / Key Responsibilities

  • Middleware & Application Architecture: Architect, operate, and support the migration of application platforms—including Messaging (Kafka, RedPanda, MQ, Pub/Sub), Service Discovery (Consul, Vault), and Data Distribution (SFTP/JScape)—to Google Cloud Platform. Manage cluster lifecycles, data replication, RBAC, and workload placement.

  • Observability & Monitoring Fabric: Design, scale, and maintain our observability backbone using tools like OpenTelemetry, Splunk, Prometheus, and Grafana. Establish and continuously improve metrics, logs, alerting strategies, SLIs, and SLOs to enable fast issue detection.

  • Incident Response & Operations: Engage with urgency in live production incidents, take ownership of minor incidents, lead post-mortems, and ensure rapid system recovery.

  • Toil Reduction & Automation: Actively identify operational toil and eliminate manual effort through code, automation, and systematic platform improvements.

  • Resiliency & Testing: Contribute to disaster recovery (DR) strategies, continuous systems resiliency testing, and present reliability improvement suggestions to the Product backlog.

  • Collaboration & Leadership: Lead technical discussions for assigned scope, present solution options, collaborate across functional teams, and mentor junior SRE colleagues.

What We're Looking For

  • Engineering & Scripting Discipline: Programming and scripting skills in high-level languages such as Python, Go, Java, or Bash to construct production-grade tooling.

  • Cloud Native & Systems Fundamentals: Proficiency with Linux-based systems, distributed systems, containerization (Kubernetes/GKE), and public cloud platforms (GCP/GCE).

  • Infrastructure as Code (IaC): Understanding of modern CI/CD patterns and IaC tools such as Terraform, Ansible, or Kubernetes Config Connector (KCC).

  • Networking & Protocols: Knowledge of core systems and networking concepts (TCP/IP, UDP, HTTP, DNS, load balancing, and messaging protocols).

  • AI & Agentic Engineering: Forward-thinking approach to automation, leveraging Generative AI and Agents (e.g., Gemini) to optimize platform operations.

  • Analytical Problem-Solving: Data-driven mindset to troubleshoot complex, non-linear system behaviors in a fast-paced, high-pressure trading ecosystem.

  • Communication & Adaptability: Strategic communication skills to translate technical requirements for cross-functional teams, coupled with an eagerness to learn independently and collaboratively.

Preferred Qualifications / Desirable

  • Observability Stack: Hands-on experience with telemetry tools such as OpenTelemetry, Splunk, Prometheus, and Grafana.

  • Agile Integration: Comfort working within Agile frameworks and collaborative software development lifecycles.

  • Certifications: GCP Professional Cloud Architect, Certified Kubernetes Administrator (CKA), or Certified Kubernetes Application Developer (CKAD).

  • Domain Expertise: Any experience in Financial Markets or other highly regulated, ultra-low latency, high-concurrency environments would be highly beneficial although not essential," 

Why CME Group?

  • Global Significance: Build technology that underpins the integrity of the world's leading derivatives marketplace.

  • Engineering Culture: Flourish in a "code-first" environment that prioritizes systematic, automated solutions over manual intervention.

  • Professional Evolution: Grow your SRE career within an organization actively transforming its approach to production engineering.

  • Competitive Package: Enjoy a robust compensation and benefits structure while working with cutting-edge tech.

Company Benefits:

  • Bonus Programme

  • Equity Programme

  • Employee Stock Purchase Plan (ESPP)

  • Private Medical and Dental coverage

  • Mental Health Benefit Programme

  • Group Pension Plan

  • Income Protection

  • Life Assurance

  • Cycle To Work

  • EV Car Benefit Scheme

  • Gym Membership

  • Family Leave

  • Education Assistance – MBA/Advanced Degree/Bachelor Degree

  • Ongoing Employee Development Training/Certification

  • Hybrid Working

#LI-RK2

#LI-Hybrid

#nijobs.com

CME Group: Where Futures are Made

CME Group is the world’s leading derivatives marketplace. But who we are goes deeper than that. Here, you can impact markets worldwide. Transform industries. And build a career by shaping tomorrow. We invest in your success and you own it – all while working alongside a team of leading experts who inspire you in ways big and small. Problem solvers, difference makers, trailblazers. Those are our people. And we’re looking for more.

At CME Group, we embrace our employees' unique experiences and skills to ensure that everyone’s perspectives are acknowledged and valued. As an equal-opportunity employer, we consider all potential employees without regard to any protected characteristic.

Important Notice: Recruitment fraud is on the rise, with scammers using misleading promises of job offers and interviews to solicit money and personal information from job seekers. CME Group adheres to established procedures designed to maintain trust, confidence and security throughout our recruitment process. Learn more here.

CME Group Belfast, Northern Ireland Office

Belfast, United Kingdom

Similar Jobs

Yesterday
Remote or Hybrid
United Kingdom
Entry level
Entry level
Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Serve as the UK technical authority for SailPoint’s Agentic AI and Non-Human Identity security products. Lead complex customer engagements from discovery through architecture, demonstrations, and Proof-of-Value execution. Act as a regional subject matter expert, deliver executive workshops and partner enablement, support industry evangelism, design secure identity architectures, and provide field feedback to influence product strategy and roadmap.
Top Skills: Aws BedrockAzure OpenaiGoogle Vertex AiIamIgaJavaJavaScriptLangchainLlamaindexPythonSailpoint Identity SecuritySpiffe/SpireWorkload Identity
Yesterday
Remote
United Kingdom
Entry level
Entry level
Cloud • Information Technology • Internet of Things • Machine Learning • Software • Cybersecurity • Infrastructure as a Service (IaaS)
Manage daily acquisition, recording, ingest, quality control, metadata, subtitles, audio, and sign-language assets for broadcast, playout, compliance, access services, and VOD. Monitor outside-source availability, record live incoming content, verify technical quality, troubleshoot issues, escalate faults, and communicate clearly with clients and internal engineering teams during time-critical transmissions.
Top Skills: Media Asset Management (Mam)Video-On-Demand (Vod)
Junior
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Perform high-volume outbound and inbound outreach (40+ calls/day) to operations leaders across Benelux, build and manage pipeline, qualify leads, collaborate with Account Executives to advance opportunities, log activity in Salesforce, represent Samsara at events, and progress through a structured ADR program toward an Account Executive role.
Top Skills: EmailLinkedInPhoneSalesforce

What you need to know about the Belfast Tech Scene

If asked to name the birthplace of the RMS Titanic, you might not say Belfast. Similarly, if asked to name Europe's leading destination for foreign direct investment in new software development, Belfast might not come to mind. Yet, both are true. The city has emerged as a tech powerhouse, recently ranked among the best in the U.K. for tech careers — especially for software developers. It also leads the U.K. with the highest percentage of software development jobs advertised.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account