Site Reliability Engineer (remote working)

Site Reliability Engineer (remote working)

Full-Time On-site
V

Working pattern: 1 day per week in the London office

Salary: Β£90,000 base + cash benefits + bonus

A leading financial services organisation is looking for an experienced Site Reliability Engineer to join a growing function within the organisation. This is a hands-on opportunity for someone who enjoys working at a technical level but also wants genuine ownership of projects and the opportunity to influence how SRE and DevOps are delivered across a large enterprise environment.

The team is continuing to develop its SRE capability, with a significant pipeline of projects focused on automation, reliability and improving the way engineering teams deliver and operate services.

The role

The successful candidate will work across Azure, Kubernetes, automation and production reliability, partnering closely with engineering and delivery teams.

Key areas of responsibility will include:

  • Improving the reliability, performance and resilience of cloud-based services
  • Reducing manual intervention across deployment and release processes
  • Automating repetitive operational tasks and reducing engineering toil
  • Helping develop SRE practices around SLOs, error budgets and reliability
  • Supporting and improving Kubernetes environments used by engineering teams
  • Working with Azure DevOps and Infrastructure as Code to improve application and environment delivery
  • Helping support the transition from Bicep towards Terraform
  • Improving the management and integration of development artefacts with wider enterprise platforms
  • Supporting production and non-production environments and taking ownership of incidents when requiredUsing monitoring and observability to identify potential reliability and performance issues
  • Working with Engineering Managers, Technical Leads and Delivery Managers to drive technical initiatives forward
  • Providing technical guidance and mentoring to less experienced engineers as the function develops

Technology skills required

The role requires strong hands-on experience with Azure and Kubernetes, alongside a solid understanding of modern DevOps and SRE practices.

  • Kubernetes and containerised environments
  • Azure DevOps and CI/CD
  • Infrastructure as Code - Terraform, Bicep or ARM
  • Azure identity, secrets and access management
  • Monitoring and observability
  • Automation and scripting
  • Microservices and cloud-native environments
  • Production support and incident management

Experience with Grafana, Azure Monitor, Log Analytics, Application Insights, PowerShell or ServiceNow would also be beneficial. Experience with both Bicep and Terraform isn't essential. The team is currently moving towards Terraform, so strong experience with either technology will be considered.

What you need

Technical capability is important, but the team is particularly interested in someone who demonstrates ownership and initiative. The successful candidate will be comfortable taking responsibility for an initiative from identifying the problem through to implementing a solution and managing the delivery themselves.

  • Strong stakeholder management is also essential. The role involves working closely with technical and non-technical stakeholders across the organisation, including Engineering Managers, Technical Leads and Delivery Managers.
  • They're also looking for someone who is naturally curious about technology and enjoys finding new ways to improve engineering practices. An interest in areas such as AI, automation and emerging technology would fit particularly well with the team's ambitions.
  • Plenty of experience working in another financial services firm

As the SRE function continues to grow, the successful candidate will have the opportunity to help establish best practice, represent the function across the wider technology community and mentor more junior engineers.

The team is at an important stage of building out its SRE capability, meaning this isn't simply a role focused on maintaining existing systems. There is already a substantial pipeline of work across SRE and DevOps, including deployment automation, Kubernetes, certificate management, reliability and operational improvements. The successful candidate will have the opportunity to shape the SRE function, reduce operational toil and establish new ways of working across a large technology organisation.

The position is remote-first, with an expectation of working from the London office approximately one day per week. The salary is Β£90k plus cash benefits and bonus.

#J-18808-Ljbffr

Site Reliability Engineer (remote working) employer: Vertus Partners

Join a leading Financial Services organisation that prioritises innovation and employee development in the heart of Bristol. With a strong focus on hybrid working, you will benefit from a collaborative work culture that encourages professional growth and offers opportunities to engage in cutting-edge data transformation initiatives. Enjoy competitive compensation and the chance to make a meaningful impact in a dynamic environment dedicated to responsible BI governance.

V

Contact Details:

Vertus Partners Recruitment Team