At a Glance
- Tasks: Ensure our SaaS platform is reliable, scalable, and operationally excellent.
- Company: Join a forward-thinking tech company focused on innovation and collaboration.
- Benefits: Enjoy competitive pay, flexible work options, and opportunities for growth.
- Other info: Be part of a diverse team committed to continuous improvement and blameless culture.
- Why this job: Make a real impact by building resilient systems and leading incident responses.
- Qualifications: Experience in Site Reliability Engineering and strong software engineering skills required.
The predicted salary is between 60000 - 80000 £ per year.
Overview
The Senior Site Reliability Engineer is responsible for the reliability, availability, scalability, and operational excellence of our Saa S platform.
This role combines software engineering, cloud infrastructure, automation, and operational leadership to build resilient systems that enable rapid product delivery.
Responsibilities
- Own the reliability and performance of production services.
- Define and operate SLIs, SLOs, and error budgets with engineering teams.
- Lead incident response and drive blameless postmortems and continuous improvement.
- Automate operational processes and reduce manual toil through engineering.
- Build and operate cloud-native platforms using Azure, Kubernetes, and Infrastructure as Code.
- Develop observability through effective monitoring, alerting, and telemetry.
- Mentor engineers and promote reliability best practices across the organisation.
Experience
- Proven experience as a Senior or experienced Site Reliability Engineer with a software engineering background.
- Ability to diagnose and make safe changes to PHP and Java or . NET applications.
- Experience operating large-scale production Saa S systems.
- Strong knowledge of SRE principles, incident management, observability, and operational excellence.
- Hands‑on experience with Azure, Kubernetes, Infrastructure as Code, and monitoring platforms such as Prometheus, Grafana, or Datadog.
- Experience influencing engineering teams and driving reliability improvements through collaboration and technical leadership.
- Key Attributes
- Passion for building reliable production systems.
- Software engineering mindset with a focus on automation.
- Strong technical judgement and ownership.
- Excellent communication during incidents and day‑to‑day collaboration.
- Commitment to blameless culture and continuous improvement.
By applying for this role, you consent to SGI contacting you by telephone regarding recruitment services, market updates, and relevant business opportunities.
We believe in equal opportunity for all and actively encourage applications from diverse backgrounds, experiences, and perspectives.
Source Group International Ltd is acting as an Employment Business in relation to this vacancy.
#J-18808-Ljbffr
Senior Site Reliability Engineer in Colchester employer: Source Technology Limited
Source Technology Limited is an exceptional employer that fosters a collaborative and innovative work culture, making it an ideal place for professionals looking to thrive in the tech industry. With a strong emphasis on employee growth and development, team members are encouraged to enhance their skills through continuous learning opportunities while working on cutting-edge cloud technologies. Located in the United Kingdom, the company offers a dynamic environment where creativity and technical expertise are valued, ensuring that employees can make a meaningful impact in their roles.