Radiant is seeking a senior Infrastructure Site Reliability Engineer for its GPU-accelerated HPC platforms. You will own reliability and performance across large-scale distributed systems, operating in a 24/7 on-call environment.
The role demands deep Linux administration, extensive HPC experience, and strong collaboration with Platform, Network, and Data Centre teams to improve observability and drive automation at scale.
#J-18808-Ljbffr