At a Glance
- Tasks: Design and develop self-service tools to enhance productivity and reduce costs.
- Company: CME Group is a leading global derivatives marketplace, shaping tomorrow's industries.
- Benefits: Enjoy hybrid working, bonus programmes, private medical coverage, and education assistance.
- Other info: Embrace a proactive culture with opportunities for ongoing development and career growth.
- Why this job: Join a diverse team, impact global markets, and innovate in a collaborative environment.
- Qualifications: 2-3 years of Python experience, knowledge of Flask/Django, and familiarity with cloud infrastructure.
The predicted salary is between 50000 - 84000 £ per year.
Job Title: Site Reliability Engineer (SRE) III – Platform Engineering & Systems ReliabilityThe Role: CME Group is seeking a Site Reliability Engineer (SRE) III to engineer reliability for our Google Cloud (GCP) infrastructure, Middleware Platform Engineering team, and core technology foundations powering our Clearing, Risk, and derivatives applications.
Be one of the first applicants, read the complete overview of the role below, then send your application for consideration.
In this role, you will help build resilient, automated systems that combine ultra-low latency with high-concurrency performance, enabling CME's product teams to innovate safely at scale.
You will work alongside senior engineers, mentor junior colleagues, engage in the dynamic operation of production systems, and assist in driving our cloud transformation.What You Will Do / Key ResponsibilitiesMiddleware & Application Architecture: Architect, operate, and support the migration of application platforms—including Messaging (Kafka, RedPanda, MQ, Pub/Sub), Service Discovery (Consul, Vault), and Data Distribution (SFTP/JScape)—to Google Cloud Platform.
Manage cluster lifecycles, data replication, RBAC, and workload placement.Observability & Monitoring Fabric: Design, scale, and maintain our observability backbone using tools like OpenTelemetry, Splunk, Prometheus, and Grafana.
Establish and continuously improve metrics, logs, alerting strategies, SLIs, and SLOs to enable fast issue detection.Incident Response & Operations: Engage with urgency in live production incidents, take ownership of minor incidents, lead post-mortems, and ensure rapid system recovery.Toil Reduction & Automation: Actively identify operational toil and eliminate manual effort through code, automation, and systematic platform improvements.Resiliency & Testing: Contribute to disaster recovery (DR) strategies, continuous systems resiliency testing, and present reliability improvement suggestions to the Product backlog.Collaboration & Leadership: Lead technical discussions for assigned scope, present solution options, collaborate across functional teams, and mentor junior SRE colleagues.What We're Looking ForEngineering & Scripting Discipline: Programming and scripting skills in high-level languages such as Python, Go, Java, or Bash to construct production-grade tooling.Cloud Native & Systems Fundamentals: Proficiency with Linux-based systems, distributed systems, containerization (Kubernetes/GKE), and public cloud platforms (GCP/GCE).Infrastructure as Code (IaC): Understanding of modern CI/CD patterns and IaC tools such as Terraform, Ansible, or Kubernetes Config Connector (KCC).Networking & Protocols: Knowledge of core systems and networking concepts (TCP/IP, UDP, HTTP, DNS, load balancing, and messaging protocols).AI & Agentic Engineering: Forward-thinking approach to automation, leveraging Generative AI and Agents (e.g., Gemini) to optimize platform operations.Analytical Problem-Solving: Data-driven mindset to troubleshoot complex, non-linear system behaviors in a fast-paced, high-pressure trading ecosystem.Communication & Adaptability: Strategic communication skills to translate technical requirements for cross-functional teams, coupled with an eagerness to learn independently and collaboratively.Preferred Qualifications / DesirableObservability Stack: Hands-on experience with telemetry tools such as OpenTelemetry, Splunk, Prometheus, and Grafana.Agile Integration: Comfort working within Agile frameworks and collaborative software development lifecycles.Certifications: GCP Professional Cloud Architect, Certified Kubernetes Administrator (CKA), or Certified Kubernetes Application Developer (CKAD).Domain Expertise: Any experience in Financial Markets or other highly regulated, ultra-low latency, high-concurrency environments would be highly beneficial although not essential," Why CME Group?Global Significance: Build technology that underpins the integrity of the world's leading derivatives marketplace.Engineering Culture: Flourish in a "code-first" environment that prioritizes systematic, automated solutions over manual intervention.Professional Evolution: Grow your SRE career within an organization actively transforming its approach to production engineering.Competitive Package: Enjoy a robust compensation and benefits structure while working with cutting-edge tech.Company Benefits:Bonus ProgrammeEquity ProgrammeEmployee Stock Purchase Plan (ESPP)Private Medical and Dental coverageMental Health Benefit ProgrammeGroup Pension PlanIncome ProtectionLife AssuranceCycle To WorkEV Car Benefit SchemeGym MembershipFamily LeaveEducation Assistance – MBA/Advanced Degree/Bachelor DegreeOngoing Employee Development Training/CertificationHybrid Working#LI-RK2#LI-Hybrid#nijobs.comCME Group: Where Futures are MadeCME Group is the world’s leading derivatives marketplace.
But who we are goes deeper than that.
Here, you can impact markets worldwide.
Transform industries.
And build a career by shaping tomorrow.
We invest in your success and you own it – all while working alongside a team of leading experts who inspire you in ways big and small.
Problem solvers, difference makers, trailblazers.
Those are our people.
And we’re looking for more.At CME Group, we embrace our employees' unique experiences and skills to ensure that everyone’s perspectives are acknowledged and valued.
As an equal-opportunity employer, we consider all potential employees without regard to any protected characteristic.Important Notice: Recruitment fraud is on the rise, with scammers using misleading promises of job offers and interviews to solicit money and personal information from job seekers.
CME Group adheres to established procedures designed to maintain trust, confidence and security throughout our recruitment process. xsabvtc
Learn more here.SummaryLocation: Belfast
- Millennium HouseType: Full time
Site Reliability Engineer III in Belfast employer: CME- Group
CME Group is an exceptional employer that prioritises innovation and employee development, particularly in the role of API-First Data & Analytics Leader. With a strong focus on cloud technologies and a collaborative work culture, employees benefit from a comprehensive bonus program, equity options, and private healthcare, all while enjoying the flexibility of hybrid working arrangements. This dynamic environment not only fosters technical growth but also encourages meaningful contributions to enhancing customer experiences.
StudySmarter Expert Advice🤫
We think this is how you could land Site Reliability Engineer III in Belfast
✨Tip Number 1
Familiarise yourself with Google Cloud Platform (GCP) and specifically Google Kubernetes Engine (GKE). Understanding how to architect and deploy microservices in GCP will give you a significant edge, as this is a key responsibility of the role.
✨Tip Number 2
Brush up on your Python skills, particularly with Flask and Django. Since the role requires developing RESTful APIs, being able to demonstrate your experience with these frameworks will be crucial during discussions.
✨Tip Number 3
Get comfortable with Infrastructure-as-Code (IaC) tools like Terraform. Being able to discuss how you've used IaC to manage cloud resources will show that you can contribute to the team's automation goals effectively.
✨Tip Number 4
Highlight any experience you have working in Agile Scrum teams. Familiarity with tools like Bitbucket and Jira will not only help you fit into the team but also demonstrate your ability to manage projects efficiently.
We think you need these skills to ace Site Reliability Engineer III in Belfast
Some tips for your application 🫡
Tailor Your CV:Make sure your CV highlights relevant experience in Python programming, automation projects, and any work with microservices or cloud platforms like GCP. Use keywords from the job description to align your skills with what CME Group is looking for.
Craft a Strong Cover Letter:In your cover letter, express your enthusiasm for the Site Reliability Engineer III role. Mention specific projects where you've designed or developed self-service tools, and how your experience aligns with their focus on automation and reliability.
Showcase Technical Skills:Clearly outline your technical skills, especially in Python frameworks like Django and Flask, as well as your familiarity with Infrastructure-as-Code using Terraform. Provide examples of how you've applied these skills in previous roles.
Highlight Team Collaboration:CME Group values collaboration within Agile teams. Include examples of how you've worked in Agile environments, used tools like Bitbucket and Jira, and contributed to team success through effective communication and problem-solving.
How to prepare for a job interview at CME- Group
✨Showcase Your Python Skills
Since the role requires approximately 2-3 years of hands-on Python programming experience, be prepared to discuss your previous projects. Highlight any automation or tooling projects you've worked on, especially those using Python Django and Flask.
✨Demonstrate Full-Stack Development Knowledge
The position involves engaging in full-stack development, so be ready to talk about your experience with both front-end and back-end technologies. Discuss any relevant projects where you delivered responsive interfaces and robust back-end services.
✨Familiarity with Infrastructure-as-Code
As the role involves managing cloud infrastructure via Infrastructure-as-Code, make sure to mention your experience with Terraform or similar tools. Be prepared to explain how you've used these tools to provision and maintain resources in a cloud environment.
✨Emphasise Your Agile Experience
CME Group values collaboration within an Agile Scrum framework. Share your experiences working in Agile teams, using tools like Bitbucket and Jira for version control and project tracking. Highlight how you contributed to sprints and managed project progress.