Senior Site Reliability Engineer (SRE) (Spain / UK)

Senior Site Reliability Engineer (SRE) (Spain / UK)

Full-Time 60000 - 80000 £ / year (est.) No working from home possible
Parser

At a Glance

  • Tasks: Design and maintain scalable systems, drive automation, and enhance operational excellence.
  • Company: Join a rapidly growing tech company with a vibrant multicultural community.
  • Benefits: Competitive salary, medical insurance, and English lessons.
  • Other info: Opportunity for mentorship and continuous improvement in a dynamic environment.
  • Why this job: Be part of a transformative team shaping the future of software development.
  • Qualifications: Proven SRE experience, Azure proficiency, and strong problem-solving skills.

The predicted salary is between 60000 - 80000 £ per year.

Senior Site Reliability Engineer (SRE)

We are seeking a highly skilled and passionate Senior Site Reliability Engineer to join our Engineering Enablement team.

This is a critical role within a large, complex, and high-impact initiative focused on deconstructing our monolithic architecture, revitalising our technology stack, and embedding quality and resilience into every stage of our development lifecycle.

You will play a pivotal role in shaping our future‑state platform, driving operational excellence, and fostering a culture of continuous improvement.

What You’ll Do

As a Senior SRE Engineer in our Engineering Enablement team, you will

  • Architect and implement reliability: design, build, and maintain highly scalable, resilient, and performant systems on Azure, focusing on our Java, Kafka, and Couchbase stack.
  • Drive modernisation: work hands‑on as part of the team spearheading the adoption of Micronaut, standardising application templates, and transitioning to managed cloud services.
  • Enhance operational excellence: develop and implement strategies for improving system observability (standardised logging, metrics, tracing), alerting, and on‑call practices.
  • Automate everything: champion automation across the software development lifecycle (SDLC), from CI/CD pipelines to infrastructure provisioning, focusing on accelerating delivery and de‑risking deployments.
  • Incident management & learning: contribute to our mature, blameless post‑incident review process, identifying root causes and implementing preventative measures to reduce incident hours.
  • Tooling & standards: develop, maintain, and drive the adoption of shared, standardised SRE tooling and best practices across engineering teams, including containerisation (e. g., Docker, Kubernetes on Azure), infrastructure as code (e. g., Terraform), and configuration management.
  • Mentorship & collaboration: provide technical leadership and mentorship to junior engineers, fostering a culture of SRE principles and operational excellence across the wider engineering organisation.
  • Strategic input: contribute to the overall technical strategy and roadmap for our SRE and platform initiatives, ensuring alignment with business objectives.

What You’ll Bring

  • Deep SRE expertise: proven experience as a Senior Site Reliability Engineer or a similar role, with a strong understanding of SRE principles (error budgets, SLOs/SLIs, toil reduction).
  • Azure cloud proficiency: extensive hands‑on experience designing, deploying, and operating highly available and scalable applications on Microsoft Azure.
  • Azure Kubernetes Service (AKS) expertise: mandatory extensive hands‑on experience with AKS for container orchestration, including deployment, scaling, monitoring, and troubleshooting.
  • Java ecosystem mastery: expert‑level proficiency with Java, including experience with modern frameworks (ideally Micronaut, Spring Boot, or similar) and JVM performance tuning.
  • Distributed systems knowledge: solid understanding and practical experience with distributed systems, microservices architecture, and associated challenges (consistency, fault tolerance).
  • Messaging & database expertise: hands‑on experience with an event streaming platform (ideally Kafka) and No SQL data storage (ideally Couchbase), including operational best practices.
  • Automation first mindset: strong scripting skills (e. g., Python, Bash) and experience with Infrastructure as Code tools (e. g., Terraform, ARM templates) and CI/CD pipelines (e. g., Azure Dev Ops, Jenkins).
  • Observability tools: experience with monitoring, logging, and alerting tools (e. g., Azure Monitor, Prometheus, Grafana, ELK Stack, Splunk).
  • Problem‑solving acumen: exceptional analytical and troubleshooting skills, with a methodical approach to diagnosing and resolving complex production issues.
  • Communication & collaboration: excellent communication skills, with the ability to articulate complex technical concepts to diverse audiences and collaborate effectively with cross‑functional teams.
  • Continuous improvement: a proactive and innovative mindset, always seeking ways to improve systems, processes, and team efficiency.

Benefits

  • The chance to join an organization with triple‑digit growth that is changing the paradigm on how software products are built.
  • The opportunity to form part of an amazing, multicultural community of tech experts.
  • A highly competitive compensation package.
  • Medical insurance.
  • English lessons.

Come and join our #Parser Community.

Follow us on Linked In.

#J-18808-Ljbffr

Senior Site Reliability Engineer (SRE) (Spain / UK) employer: Parser

Parser is an exceptional employer that champions a culture of innovation and collaboration, offering Senior Java Engineers the chance to work on impactful projects in a fast-growing, multicultural environment. With a hybrid working model based in the UK, employees benefit from competitive compensation, medical insurance, and opportunities for continuous learning alongside top-tier specialists. Join us to take ownership of your work and contribute to meaningful digital transformations while enjoying a flexible work-life balance.

Parser

Contact Details:

Parser Recruitment Team

We think you need these skills to ace Senior Site Reliability Engineer (SRE) (Spain / UK)

Site Reliability Engineering (SRE)
Azure Cloud Proficiency
Azure Kubernetes Service (AKS)
Java Ecosystem Mastery
Distributed Systems Knowledge
Messaging & Database Expertise (Kafka, Couchbase)
Automation Skills (Python, Bash)