Jobgether is seeking a Senior Site Reliability Engineer to own compute node reliability for a large-scale AI and cloud platform in the United Kingdom. You will operate close to Linux OS, hypervisor, and node-level services, focusing on Linux systems engineering, virtualization, containers, observability, and production reliability.
The role involves investigating complex system issues, shaping reliability practices, and collaborating with platform, kernel, and infrastructure teams to improve
#J-18808-Ljbffr
Senior SRE: Compute Node Platform for AI & Cloud employer: Jobgether
As a Senior ML Engineer at our innovative healthcare-focused company in the UK, you will be part of a dynamic team dedicated to transforming clinical environments through cutting-edge technology. We offer a collaborative work culture that fosters creativity and professional growth, alongside competitive benefits and opportunities for continuous learning in the rapidly evolving field of machine learning. Join us to make a meaningful impact on healthcare while enjoying a supportive environment that values your contributions.