At a Glance
- Tasks: Design and maintain scalable AWS infrastructure while optimising Kubernetes environments.
- Company: Join a fast-growing SaaS company in Central London with a hybrid work model.
- Benefits: Up to £85,000 salary, bonus, and great benefits including professional development.
- Other info: Opportunities for career growth as the company scales and matures its SRE practices.
- Why this job: Take ownership of platform reliability and influence engineering decisions in a dynamic environment.
- Qualifications: Experience with AWS, Kubernetes, Terraform, and a passion for automation and reliability.
Up to £85,000 + Benefits
Central London – Hybrid (2/3 days a week in the office)
We're partnering with a fast-growing SaaS company that's going through an exciting period of growth and investing heavily in its engineering and platform capabilities. They're looking for an experienced Site Reliability Engineer (SRE) to join the team and play a key role in building highly reliable, scalable, and observable infrastructure. This is a hands-on role focused on AWS, Kubernetes, Terraform, observability, monitoring, and automation, working closely with software engineering teams to improve platform reliability and developer experience. You'll have genuine ownership and the opportunity to influence how the platform evolves as the business continues to scale.
What You'll Be Doing
- Design, build, and maintain highly available and scalable AWS infrastructure
- Manage and optimise Kubernetes environments and containerised workloads
- Build and maintain infrastructure using Terraform and Infrastructure as Code principles
- Develop and optimise CI/CD pipelines using GitHub Actions
- Build and improve comprehensive monitoring and observability across the platform
- Implement and maintain effective logging, metrics, tracing, alerting, and dashboards
- Define and improve SLIs, SLOs, and reliability metrics
- Proactively identify and resolve performance, availability, and reliability issues
- Lead and contribute to incident response, troubleshooting, and root cause analysis
- Automate operational processes and eliminate repetitive manual tasks
- Work closely with software engineers to improve deployment processes, system reliability, and developer experience
- Help improve platform resilience, scalability, and disaster recovery capabilities
- Contribute to capacity planning and performance optimisation as the platform scales
- Establish and champion SRE best practices across the wider engineering function
What We're Looking For
- Proven commercial experience working as an SRE, DevOps Engineer, Platform Engineer, or similar
- Strong hands-on experience with AWS
- Strong experience working with Kubernetes
- Excellent experience with Terraform and Infrastructure as Code
- Strong experience building and managing GitHub Actions CI/CD pipelines
- Solid experience with monitoring and observability tooling
- Strong understanding of metrics, logging, tracing, alerting, and system health
- Experience troubleshooting complex production environments
- Understanding of SLIs, SLOs, SLAs, and error budgets
- Experience with incident management and root cause analysis
- Good understanding of cloud networking, security, and infrastructure fundamentals
- Strong scripting/automation skills
- A strong understanding of reliability, scalability, performance, and availability
- Excellent communication skills and the ability to work closely with software engineering teams
- A proactive mindset and genuine passion for automation and continuous improvement
Don't Tick Every Box?
That's okay. The company is open to speaking with engineers who may not have experience across every technology listed above. If you have strong foundations in AWS, Kubernetes, Terraform, and cloud infrastructure, along with a genuine interest in reliability and observability, we'd still love to hear from you.
Why Join?
- Join a fast-growing SaaS company at an exciting stage of its journey
- Work with a modern AWS and Kubernetes environment
- Take ownership of reliability, automation, and platform performance
- Work with modern observability and monitoring technologies
- Have genuine influence over engineering and platform decisions
- Work closely with talented software engineering teams
- Clear opportunities to progress as the business continues to scale
- Help shape and mature the company's SRE practices
- Hybrid working from Central London, 2/3 days per week
If you're an experienced SRE, Platform Engineer or DevOps Engineer who enjoys solving complex reliability challenges and wants to have a real impact within a rapidly growing SaaS business, we'd love to hear from you. Apply now or get in touch for a confidential conversation.
Site Reliability Engineer in London employer: ReVybe IT Recruitment Limited
REVYBE IT RECRUITMENT LIMITED is an exceptional employer located in the vibrant City Centre of Manchester, offering a dynamic work culture that fosters collaboration and innovation. Employees benefit from competitive salaries, comprehensive benefits, and opportunities for professional growth within a forward-thinking Fintech environment that prioritises security integration from day one. Join a team where your contributions are valued, and you can make a meaningful impact in the world of DevSecOps.
Contact Details:
ReVybe IT Recruitment Limited Recruitment Team
StudySmarter Expert Advice🤫
We think this is how you could land Site Reliability Engineer in London
✨Join Local Tech Meetups
Get out there and mingle with fellow developers by joining local tech meetups. It’s a fantastic way to meet people who might be working at ReVybe IT Recruitment Limited or know someone who does. Plus, you can pick up some trendy tech skills and trends while you're at it!
✨Contribute to Open Source Projects
Show off your coding chops by jumping into open-source projects. Not only does this give you practical experience, but it also gets you noticed in the dev community. You'll create a killer portfolio that speaks volumes about your skills to ReVybe IT Recruitment Limited.
✨Tap into Online Developer Communities
Don’t underestimate the power of online developer communities like GitHub, Stack Overflow, and even Reddit. Participate in discussions, share your projects, and build your visibility. We can often find opportunities through these channels that can lead to a full-time gig at companies like ReVybe IT Recruitment Limited.
✨Explore Job Boards Specifically for Tech Roles
Keep your eyes peeled on job boards that focus on tech roles. Sites like TechCareers or Stack Overflow Jobs can often have listings for companies like ReVybe IT Recruitment Limited that might not show up on broader job sites. Make it a habit to check these regularly, and don’t hesitate to apply directly through our website!
We think you need these skills to ace Site Reliability Engineer in London
Some tips for your application 🫡
Show off your coding skills:When applying for a software engineering role, it's super important to showcase your coding skills. Make sure your CV includes your tech stack, any relevant programming languages you’re comfortable with, and examples of projects you've worked on. If you have a GitHub profile, link it up! We love to see code in action.
Tailor your portfolio:For a full-time role, we’d expect to see some solid examples of your work in your portfolio. Make sure to include at least two or three projects that highlight your problem-solving skills and your ability to work with different technologies. Focus on the projects that are most relevant to the position at ReVybe IT Recruitment Limited.
Craft a killer cover letter:Your cover letter is your chance to stand out—make it personal! Explain why you want to work at ReVybe IT Recruitment Limited and how your skills align with the role. Show us your passion for software development. We dig enthusiastic candidates who understand the value of collaboration and continuous learning!
Be clear and concise:When it comes to writing your CV and cover letter, clarity is key. Avoid jargon that could confuse us and stick to simple, direct language. Highlight your achievements with quantifiable results where possible, and keep everything easy to read. A well-organised application goes a long way!
How to prepare for a job interview at ReVybe IT Recruitment Limited
✨Brush Up on Your Coding Skills
For a full-time software engineering role, it's crucial that we stay sharp with our coding abilities. Expect technical questions that might involve solving problems on the spot or discussing algorithms. Practise on platforms like LeetCode or HackerRank to get comfortable with the types of questions that often come up.
✨Know Your Tools and Frameworks
Make sure we’re well-acquainted with the tools and technologies listed in the job description. Familiarise ourselves with any specific frameworks or programming languages mentioned. If ReVybe IT Recruitment Limited uses React or Node.js, for instance, be ready to discuss how we’ve used them in previous projects or coursework.
✨Showcase Your Projects
Bring along a portfolio that highlights our best work. This could be code samples, GitHub repositories, or any side projects we’ve built. Make sure we can talk through our thought process for each project, especially the challenges we faced and how we solved them—this shows our problem-solving skills in action.
✨Prepare for Behavioural Questions
While technical skills are key, full-time positions also require cultural fit. Be ready to discuss our previous experiences and how we handle teamwork, conflict, and deadlines. Brush up on the STAR method—Situation, Task, Action, Result—to clearly articulate our past experiences when discussing how we've contributed to a team.