Senior Site Reliability Engineer - MLOps & HPC Platforms in London

Senior Site Reliability Engineer - MLOps & HPC Platforms in London

London Full-Time 63000 - 77000 £ / year (est.) No working from home possible
Roche

At a Glance

  • Tasks: Design resilient cloud platforms for MLOps and HPC workloads to accelerate medicine discovery.
  • Company: Join Roche, a leader in healthcare innovation, based in vibrant London.
  • Benefits: Enjoy competitive salary, health perks, remote work options, and career development opportunities.
  • Other info: Be part of a dynamic team driving transformative change in the medical field.
  • Why this job: Make a real impact in healthcare by building reliable systems on a global scale.
  • Qualifications: Experience in site reliability engineering and cloud platforms like AWS, Azure, or GCP.

The predicted salary is between 63000 - 77000 £ per year.

Roche in London is seeking a Senior Site Reliability Engineer within the Computational Sciences Center of Excellence to design resilient, cloud‑based platforms for MLOps and HPC workloads at global scale, enabling faster discovery of transformative medicines.

You will implement Infrastructure as Code with Terraform, Pulumi or CloudFormation, build autoscaling and disaster recovery, drive deep observability, and lead a team of engineers delivering reliable systems across AWS, Azure and GCP.

Senior Site Reliability Engineer - MLOps & HPC Platforms in London employer: Roche

Roche is an exceptional employer, offering a dynamic work environment in the heart of pharmaceutical innovation. With a strong commitment to employee growth, we provide opportunities for mentorship and professional development, fostering a culture of curiosity and collaboration. Our location enables access to cutting-edge resources and a diverse team, empowering you to make a meaningful impact on global healthcare challenges.

Roche

Contact Details:

Roche Recruitment Team

We think you need these skills to ace Senior Site Reliability Engineer - MLOps & HPC Platforms in London

Site Reliability Engineering
MLOps
HPC Platforms
Infrastructure as Code
Terraform
Pulumi
CloudFormation