Site Reliability Engineer

Site Reliability Engineer

Full-Time 36000 - 60000 £ / year (est.) No working from home possible
A

At a Glance

  • Tasks: Shape and drive reliable, observable systems while collaborating with development teams.
  • Company: Leading gaming software provider transforming the future of gaming.
  • Benefits: Competitive salary, flexible working, and opportunities for professional growth.
  • Other info: Dynamic team environment with a focus on innovation and collaboration.
  • Why this job: Make a real impact on system reliability and customer experiences in gaming.
  • Qualifications: 3+ years in SRE or DevOps, strong Kubernetes and scripting skills.

The predicted salary is between 36000 - 60000 £ per year.

About Us

Amelco Ltd are a leading gaming and gambling solution software provider with a strong presence in the USA, UK, and Europe. Through partnerships with global gaming companies, we build cutting-edge technical platforms across sportsbooks, lottery, casino, virtual gaming, and financial trading. Our vision is to shape the future of gaming by transforming operations into intelligent, data-driven solutions that deliver exceptional customer experiences and create sustainable value for all stakeholders. We believe in teamwork, knowledge sharing, and transparency with accountability.

The Role

We’re looking for a Site Reliability Engineer (SRE) to help shape and drive how we build and operate reliable, observable, and cost-efficient systems. You’ll work closely with our development, platform, and incident management teams to define what “reliable” means in measurable terms — and build the tooling and processes to achieve it. Your work will directly influence the speed, stability, and scalability of our platform.

Key Responsibilities

  • Partner with development teams to define and manage SLOs/SLIs, and use error budgets to guide engineering decisions.
  • Enhance observability — ensuring metrics, logs, and tracing are in place to detect and fix issues proactively.
  • Lead cost optimisation initiatives: monitor spend, rightsize workloads, tune autoscaling, and drive efficient infrastructure usage.
  • Strengthen production readiness with pre-deployment checks, post-release validation, and robust platform guardrails.
  • Introduce and run chaos engineering experiments to improve system resilience.
  • Automate operational processes to reduce manual intervention across the stack.
  • Contribute to major incident response, providing engineering expertise.
  • Collaborate cross-functionally to raise the bar on platform stability, security, and performance.

Required Skills & Experience

  • 3+ years in SRE, Platform, or DevOps roles.
  • Strong operational experience with Kubernetes (on-prem and AWS EKS).
  • Proven track record defining and working with SLOs/SLIs in production environments.
  • Deep understanding of observability (metrics, logging, tracing, telemetry).
  • Experience with Terraform and GitOps workflows.
  • Proficient scripting skills in Python, Bash, or Go.
  • Demonstrated ability to balance cost efficiency and system reliability.
  • Excellent communication skills and ability to work across multiple teams.
  • Experience running chaos engineering experiments.
  • Exposure to high-throughput, low-latency systems.
  • Experience with FinOps or cost management practices.
  • AWS Certifications (e.g., Solutions Architect, DevOps Engineer).

Site Reliability Engineer employer: Amelco Limited

Amelco Limited is an excellent employer, offering a dynamic work environment in the vibrant area of Shoreditch, London. With a strong focus on employee growth, we provide opportunities for professional development and travel, alongside a competitive pension scheme. Our collaborative culture encourages innovation and problem-solving, making it a rewarding place for skilled database administrators to thrive.

A

Contact Details:

Amelco Limited Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Site Reliability Engineer

Tip Number 1

Network like a pro! Reach out to folks in the industry, attend meetups, and connect with potential colleagues on LinkedIn. You never know who might have the inside scoop on job openings or can put in a good word for you.

Tip Number 2

Show off your skills! Create a portfolio or GitHub repository showcasing your projects, especially those related to SRE, Kubernetes, or automation. This gives employers a tangible look at what you can do and sets you apart from the crowd.

Tip Number 3

Prepare for interviews by brushing up on your technical knowledge and soft skills. Practice common SRE scenarios and be ready to discuss how you've tackled challenges in the past. Confidence is key!

Tip Number 4

Don't forget to apply through our website! We love seeing candidates who are genuinely interested in joining our team. Tailor your application to highlight how your experience aligns with our mission of delivering exceptional customer experiences.

We think you need these skills to ace Site Reliability Engineer

Site Reliability Engineering (SRE)
Kubernetes
AWS EKS
SLOs/SLIs management
Observability (metrics, logging, tracing, telemetry)
Terraform
GitOps workflows

Some tips for your application 🫡

Show Your Passion for Reliability:When writing your application, let us see your enthusiasm for building reliable systems. Share specific examples of how you've improved system reliability in past roles, especially if you’ve worked with SLOs and SLIs.

Be Clear and Concise:We appreciate clarity! Make sure your application is easy to read and straight to the point. Use bullet points where necessary to highlight your skills and experiences, especially those related to Kubernetes and observability.

Tailor Your Application:Don’t just send a generic application. Tailor it to our job description by mentioning relevant experiences that align with our needs, like your work with Terraform or chaos engineering experiments. Show us why you’re the perfect fit!

Apply Through Our Website:We encourage you to apply through our website for the best chance of getting noticed. It’s super easy, and you’ll be able to showcase your skills directly to us. We can’t wait to see what you bring to the table!

How to prepare for a job interview at Amelco Limited

Know Your SLOs and SLIs

Make sure you understand the concepts of Service Level Objectives (SLOs) and Service Level Indicators (SLIs) inside out. Be ready to discuss how you've defined and managed these in your previous roles, as this will show your practical experience and understanding of reliability.

Showcase Your Observability Skills

Prepare examples of how you've enhanced observability in past projects. Talk about the tools you've used for metrics, logging, and tracing, and how they helped you proactively detect and fix issues. This will demonstrate your hands-on experience with observability.

Discuss Cost Optimisation Strategies

Be ready to share specific instances where you've led cost optimisation initiatives. Discuss how you monitored spend, right-sized workloads, and tuned autoscaling. This shows that you can balance cost efficiency with system reliability, which is crucial for the role.

Emphasise Collaboration and Communication

Since the role involves working closely with various teams, highlight your communication skills and any cross-functional projects you've been part of. Share how you’ve collaborated to improve platform stability and performance, as teamwork is key at Amelco Ltd.