CME Group Logo

CME Group

Site Reliability Engineer III

Posted 9 Days Ago
Remote or Hybrid
Hiring Remotely in Whitehouse, Belfast, Northern Ireland, GBR
Entry level
Remote or Hybrid
Hiring Remotely in Whitehouse, Belfast, Northern Ireland, GBR
Entry level
Engineer and operate reliable GCP infrastructure and middleware platforms supporting high-concurrency, ultra-low-latency trading applications. Responsibilities include migrating messaging, service discovery, and data distribution systems; maintaining observability, SLIs, and SLOs; responding to production incidents; reducing toil through automation; supporting disaster recovery and resiliency testing; and mentoring junior engineers.
The summary above was generated by AI

Job Title: Site Reliability Engineer (SRE) III – Platform Engineering & Systems Reliability

The Role: CME Group is seeking a Site Reliability Engineer (SRE) III to engineer reliability for our Google Cloud (GCP) infrastructure, Middleware Platform Engineering team, and core technology foundations powering our Clearing, Risk, and derivatives applications. In this role, you will help build resilient, automated systems that combine ultra-low latency with high-concurrency performance, enabling CME's product teams to innovate safely at scale. You will work alongside senior engineers, mentor junior colleagues, engage in the dynamic operation of production systems, and assist in driving our cloud transformation.

What You Will Do / Key Responsibilities

  • Middleware & Application Architecture: Architect, operate, and support the migration of application platforms—including Messaging (Kafka, RedPanda, MQ, Pub/Sub), Service Discovery (Consul, Vault), and Data Distribution (SFTP/JScape)—to Google Cloud Platform. Manage cluster lifecycles, data replication, RBAC, and workload placement.

  • Observability & Monitoring Fabric: Design, scale, and maintain our observability backbone using tools like OpenTelemetry, Splunk, Prometheus, and Grafana. Establish and continuously improve metrics, logs, alerting strategies, SLIs, and SLOs to enable fast issue detection.

  • Incident Response & Operations: Engage with urgency in live production incidents, take ownership of minor incidents, lead post-mortems, and ensure rapid system recovery.

  • Toil Reduction & Automation: Actively identify operational toil and eliminate manual effort through code, automation, and systematic platform improvements.

  • Resiliency & Testing: Contribute to disaster recovery (DR) strategies, continuous systems resiliency testing, and present reliability improvement suggestions to the Product backlog.

  • Collaboration & Leadership: Lead technical discussions for assigned scope, present solution options, collaborate across functional teams, and mentor junior SRE colleagues.

What We're Looking For

  • Engineering & Scripting Discipline: Programming and scripting skills in high-level languages such as Python, Go, Java, or Bash to construct production-grade tooling.

  • Cloud Native & Systems Fundamentals: Proficiency with Linux-based systems, distributed systems, containerization (Kubernetes/GKE), and public cloud platforms (GCP/GCE).

  • Infrastructure as Code (IaC): Understanding of modern CI/CD patterns and IaC tools such as Terraform, Ansible, or Kubernetes Config Connector (KCC).

  • Networking & Protocols: Knowledge of core systems and networking concepts (TCP/IP, UDP, HTTP, DNS, load balancing, and messaging protocols).

  • AI & Agentic Engineering: Forward-thinking approach to automation, leveraging Generative AI and Agents (e.g., Gemini) to optimize platform operations.

  • Analytical Problem-Solving: Data-driven mindset to troubleshoot complex, non-linear system behaviors in a fast-paced, high-pressure trading ecosystem.

  • Communication & Adaptability: Strategic communication skills to translate technical requirements for cross-functional teams, coupled with an eagerness to learn independently and collaboratively.

Preferred Qualifications / Desirable

  • Observability Stack: Hands-on experience with telemetry tools such as OpenTelemetry, Splunk, Prometheus, and Grafana.

  • Agile Integration: Comfort working within Agile frameworks and collaborative software development lifecycles.

  • Certifications: GCP Professional Cloud Architect, Certified Kubernetes Administrator (CKA), or Certified Kubernetes Application Developer (CKAD).

  • Domain Expertise: Any experience in Financial Markets or other highly regulated, ultra-low latency, high-concurrency environments would be highly beneficial although not essential," 

Why CME Group?

  • Global Significance: Build technology that underpins the integrity of the world's leading derivatives marketplace.

  • Engineering Culture: Flourish in a "code-first" environment that prioritizes systematic, automated solutions over manual intervention.

  • Professional Evolution: Grow your SRE career within an organization actively transforming its approach to production engineering.

  • Competitive Package: Enjoy a robust compensation and benefits structure while working with cutting-edge tech.

Company Benefits:

  • Bonus Programme

  • Equity Programme

  • Employee Stock Purchase Plan (ESPP)

  • Private Medical and Dental coverage

  • Mental Health Benefit Programme

  • Group Pension Plan

  • Income Protection

  • Life Assurance

  • Cycle To Work

  • EV Car Benefit Scheme

  • Gym Membership

  • Family Leave

  • Education Assistance – MBA/Advanced Degree/Bachelor Degree

  • Ongoing Employee Development Training/Certification

  • Hybrid Working

#LI-RK2

#LI-Hybrid

#nijobs.com

CME Group: Where Futures are Made

CME Group is the world’s leading derivatives marketplace. But who we are goes deeper than that. Here, you can impact markets worldwide. Transform industries. And build a career by shaping tomorrow. We invest in your success and you own it – all while working alongside a team of leading experts who inspire you in ways big and small. Problem solvers, difference makers, trailblazers. Those are our people. And we’re looking for more.

At CME Group, we embrace our employees' unique experiences and skills to ensure that everyone’s perspectives are acknowledged and valued. As an equal-opportunity employer, we consider all potential employees without regard to any protected characteristic.

Important Notice: Recruitment fraud is on the rise, with scammers using misleading promises of job offers and interviews to solicit money and personal information from job seekers. CME Group adheres to established procedures designed to maintain trust, confidence and security throughout our recruitment process. Learn more here.

CME Group London, England Office

3A One New Change, London, United Kingdom, EC4M 9AF

Similar Jobs

Yesterday
Easy Apply
Remote
United Kingdom
Easy Apply
Senior level
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Conduct advanced application security research across GitLab’s DevSecOps and AI-powered platforms. Identify and exploit novel, systemic, and chained vulnerabilities; develop proof-of-concept exploits, testing methodologies, and research automation; assess AI attack vectors and open-source dependencies; drive remediation and security improvements; advise engineering teams, contribute to roadmaps, mentor researchers, and share findings with the security community.
Top Skills: Duo Agent PlatformGitlabGitlab DuoGoPythonRubyRustTypescript
Yesterday
Easy Apply
Remote
United Kingdom
Easy Apply
Expert/Leader
Expert/Leader
Cloud • Security • Software • Cybersecurity • Automation
Conducts advanced application security research across GitLab’s DevSecOps and AI-powered platforms. Responsibilities include discovering and exploiting systemic vulnerabilities, penetration testing, developing proof-of-concept exploits, researching AI and agentic security threats, building automated research tooling, coordinating remediation, mentoring security professionals, and communicating findings to engineering and security communities.
Top Skills: Ai FrameworksDevsecopsDuo Agent PlatformGitlabGitlab Duo ChatGoPythonRubyRustTypescript
2 Days Ago
Easy Apply
Remote or Hybrid
UK
Easy Apply
Expert/Leader
Expert/Leader
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Lead Samsara’s vulnerability management and application security programs across cloud, firmware, IoT, and corporate systems. Define technical strategy, automate vulnerability detection and response, reduce remediation time, guide engineering teams, mentor security engineers, investigate critical vulnerabilities, and support incident response. The role requires strong AWS, programming, vulnerability management, application security testing, AI/LLM, and security automation expertise.
Top Skills: Ai/LlmAWSAws LambdaCC++Ci/CdCvssDastEpssFedrampFirmwareGoIotJavaScriptPythonSastScaSemgrepTinesWiz

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account