Senior SRE - UK: Lead Incidents & Reliability in London

Senior SRE - UK: Lead Incidents & Reliability in London

London Full-Time 60000 - 80000 £ / year (est.) No working from home possible
V

At a Glance

  • Tasks: Lead incident response and enhance production reliability in a fast-paced environment.
  • Company: Heidi, a dynamic tech company based in London.
  • Benefits: Competitive salary, on-call bonuses, and opportunities for professional growth.
  • Other info: Emphasises blameless post-mortems and continuous improvement.
  • Why this job: Join a passionate team and make a real impact on system reliability.
  • Qualifications: Experience with Kubernetes, AWS, and a collaborative mindset.

The predicted salary is between 60000 - 80000 £ per year.

Heidi in London (on-site) is seeking a Senior Site Reliability Engineer to join the Platform/SRE team, owning production reliability and incident response.

You’ll work on on‑call rotations, improving alerting, automation, and deployments, with hands‑on experience in Kubernetes and AWS.

Candidates should be comfortable in a fast‑paced, ops‑heavy environment and collaborate across product teams.

The role emphasizes reliability practices, blameless post‑mortems, and practical improvements to reduce

#J-18808-Ljbffr

Senior SRE - UK: Lead Incidents & Reliability in London employer: Voice AI Space

HappyRobot is an exceptional employer that fosters a dynamic and innovative work culture in the heart of London. As a fast-growing AI startup, we offer our employees meaningful opportunities for professional growth and development, alongside competitive benefits and a collaborative environment that encourages ownership and creativity in legal practices. Join us to be part of a team that values your contributions and supports your career journey in a thriving industry.

V

Contact Details:

Voice AI Space Recruitment Team

We think you need these skills to ace Senior SRE - UK: Lead Incidents & Reliability in London

Incident Response
Production Reliability
On-Call Rotations
Alerting Improvement
Automation
Deployments
Kubernetes