Senior HPC Infra SRE: GPU Compute, 24/7 Reliability

Senior HPC Infra SRE: GPU Compute, 24/7 Reliability

Full-Time No working from home possible
R

Radiant is seeking a senior Infrastructure Site Reliability Engineer for its GPU-accelerated HPC platforms. You will own reliability and performance across large-scale distributed systems, operating in a 24/7 on-call environment.

The role demands deep Linux administration, extensive HPC experience, and strong collaboration with Platform, Network, and Data Centre teams to improve observability and drive automation at scale.

#J-18808-Ljbffr

R

Contact Details:

Radiant Recruitment Team