HPC Fleet Reliability Engineer - GPU Clusters

HPC Fleet Reliability Engineer - GPU Clusters

Full-Time 60000 - 75000 £ / year (est.) No working from home possible
Dormont Manufacturing Co

At a Glance

  • Tasks: Manage and optimise supercomputing clusters for peak performance.
  • Company: Dormont Manufacturing Co, a leader in tech innovation.
  • Benefits: Medical and dental insurance, pension contributions, and more.
  • Other info: Exciting opportunities for growth in a dynamic work environment.
  • Why this job: Join a fast-paced team and make an impact in cutting-edge technology.
  • Qualifications: Bachelor’s degree and 2 years of experience in data centre infrastructure.

The predicted salary is between 60000 - 75000 £ per year.

Dormont Manufacturing Co is seeking a Fleet Reliability Operations team member to oversee the management and uptime of supercomputing clusters.

The successful candidate will configure and troubleshoot issues in a fast-paced environment, ensuring optimal performance of our systems and infrastructure.

This role requires a bachelor’s degree and at least 2 years of experience in troubleshooting and maintaining data center infrastructure, ideally in a Linux environment.

The position offers various benefits including medical and dental insurance, and pension contributions.

#J-18808-Ljbffr

HPC Fleet Reliability Engineer - GPU Clusters employer: Dormont Manufacturing Co

Dormont Manufacturing Co is an exceptional employer, offering a dynamic work environment that fosters innovation and collaboration. With a strong commitment to employee growth, the company provides extensive training opportunities and a comprehensive benefits package, all while supporting a flexible hybrid work model that enhances work-life balance. Join us in shaping the future of lottery solutions in a role that promises meaningful impact and professional development.

Dormont Manufacturing Co

Contact Details:

Dormont Manufacturing Co Recruitment Team

We think you need these skills to ace HPC Fleet Reliability Engineer - GPU Clusters

Supercomputing Clusters Management
Troubleshooting Skills
Data Center Infrastructure Maintenance
Linux Environment Proficiency
System Configuration
Performance Optimisation
Fast-Paced Environment Adaptability