Systems Engineer (Monitoring) in London

Systems Engineer (Monitoring) in London

London Full-Time 59400 - 72600 Β£ / year (est.) No working from home possible
Oxford Knight

At a Glance

  • Tasks: Join a top trading firm to optimise GPU systems and automate large-scale infrastructure.
  • Company: Leading algorithmic trading firm with a cutting-edge research environment.
  • Benefits: Competitive salary, innovative projects, and a diverse team of experts.
  • Other info: Dynamic culture that values fresh ideas from all backgrounds.
  • Why this job: Make a real impact in high-performance computing and AI while collaborating with industry leaders.
  • Qualifications: 5+ years in Linux systems engineering, GPU optimisation, and Python scripting.

The predicted salary is between 59400 - 72600 Β£ per year.

One of the world's top algorithmic trading firms, our client is looking for GPU Systems Engineers to help scale and evolve their exceptionally sophisticated HPC/AI research environment. Joining the Research and Development team, you will collaborate with experts responsible for the compute, storage, operating systems, and automation tools that enable trading and research to run 24/7 across the globe. They design, grow, and operate infrastructure at a large scale, including triple-digit petabyte-scale storage and massive CPU and GPU clusters in globally distributed data centers. As such, this is a high-impact role with broad scope, from HPC/AI cluster design and performance tuning, to troubleshooting and automation for thousands of nodes.

  • Identify and resolve GPU workloads' performance bottlenecks across compute, storage, and networking layers
  • Automate system deployment, monitoring, and troubleshooting across thousands of nodes
  • Collaborate with research and engineering teams to support evolving workloads
  • Own critical infrastructure projects - from concept to implementation and support
  • Test and deploy new hardware and software, and partner with vendors to resolve complex issues

5+ years of experience in large-scale Linux systems engineering in HPC, AI or distributed infrastructure roles

  • Extensive experience in Linux system installation, performance tuning, and troubleshooting
  • Deep knowledge around GPU optimization and performance
  • Proficiency in Python scripting and automation frameworks
  • CUDA or C/C++ experience is a plus
  • Experience with NVIDIA technologies beyond CUDA, such as NCCL, GPUDirect RDMA, and NVLink
  • Familiarity with configuration management tools
  • Comfortable diagnosing complex system issues at the hardware, OS, and network levels

This fund brings a scientific approach to trading financial products. They've built one of the world's most sophisticated computing environments for research and development, and their researchers are at the forefront of innovation in the world of algorithmic trading. Colleagues come from all sorts of backgrounds: mathematics, computer science, statistics, physics, and engineering. A community of self-starters who are motivated by the excitement of being at the cutting edge of automated trading, and their culture celebrates great ideas whether they come from veterans or new hires.

Systems Engineer (Monitoring) in London employer: Oxford Knight

As one of the world's leading algorithmic trading firms, we offer an exceptional work environment where innovation thrives. Our friendly and informal culture fosters collaboration across teams, allowing you to tackle complex challenges with cutting-edge technologies while enjoying a market-leading salary and generous benefits. Join us to make a significant impact in the financial sector and advance your career in a role that values your contributions and expertise.

Oxford Knight

Contact Details:

Oxford Knight Recruitment Team

We think you need these skills to ace Systems Engineer (Monitoring) in London

Linux Systems Engineering
HPC (High-Performance Computing)
AI (Artificial Intelligence)
GPU Optimization
Performance Tuning
Troubleshooting
Python Scripting