Location: London, onsite 3 days per week (Sheffield as an alternative) Rate: £tbd/day inside IR35 Duration: 6 months+ Are you a Senior Observability Engineer / SRE Lead, with demonstrable experience of assessing and defining observability and monitoring roadmaps within enterprise scale environments? If so, apply now for this new contract opportunity. The Lead Observability Engineer / SRE Lead will be required to assess a complex hybrid estate, understand how services, platforms, infrastructure and networks should be monitored, and work across multiple internal teams, partners and suppliers to build a consolidated view of existing telemetry, monitoring and alerting capabilities. As well as short term tactical objectives, the role will also be focussed on longer term strategic ones. Responsibilities of the Lead Observability Engineer / SRE Lead will be to: Discover and document existing telemetry sources, monitoring tools, dashboards and ownership Work with technical teams and suppliers to gain access to telemetry Deliver meaningful dashboards and service health views Identify gaps in telemetry, monitoring and alerting coverage and implement pragmatic improvements Define health indicators for critical business journeys Introduce modern observability practices where practical, including SLIs, SLOs and a roadmap towards burn-rate alerting Develop a roadmap for OpenTelemetry adoption Assess options for a centralised telemetry platform, including Grafana Cloud Evaluate tooling rationalisation opportunities, operating costs and telemetry economics Define an observability target architecture, standards and implementation roadmap The successful Lead Observability Engineer / SRE Lead will demonstrate the following: Proven experience leading enterprise-scale observability initiatives Strong hands-on expertise with Grafana, Grafana Cloud, OpenTelemetry and modern telemetry pipelines Deep understanding of metrics, logs, traces, distributed tracing, alerting, SLIs, SLOs and error-budget concepts Experience designing observability solutions across cloud PaaS, IaaS, on-premises, legacy and third-party hosted platforms Strong knowledge of Azure observability tooling, including Azure Monitor, Log Analytics and Application InsightsIf this sounds like you, please apply to find out more. Lead Observability Engineer / SRE Lead / Lead Site Reliability Engineer
Lead Observability Engineer in London employer: TRIA
Tria is an exceptional employer that fosters a dynamic work culture, offering hybrid working arrangements that promote work-life balance. With a strong focus on employee growth and development, Tria provides opportunities for professional advancement while leading transformative projects in the North West. Join a team that values innovation and collaboration, ensuring you play a pivotal role in shaping the future of enterprise infrastructure.