Senior Solution Engineer – GPU & AI Infrastructure

Senior Solution Engineer – GPU & AI Infrastructure

Full-Time 72000 - 88000 £ / year (est.) Home office (partial)
Civo Ltd

At a Glance

  • Tasks: Design cutting-edge GPU clusters for AI and HPC projects, bridging tech and business needs.
  • Company: Join Civo, a pioneering neocloud provider revolutionising AI infrastructure.
  • Benefits: Enjoy a competitive salary, flexible remote work, and unlimited holiday.
  • Other info: Be part of a dynamic team in a collaborative and inclusive culture.
  • Why this job: Make a real impact in the fast-paced world of AI and high-performance computing.
  • Qualifications: 5+ years in solution architecture with expertise in NVIDIA GPU technologies.

The predicted salary is between 72000 - 88000 £ per year.

Senior Solution Engineer – GPU & AI Infrastructure

About Civo

Civo is a high-performance neocloud provider purpose-built for the demands of modern AI, high-performance computing (HPC), and cloud-native infrastructure.

We eliminate legacy cloud overhead to deliver ultra-low-latency compute, bare-metal GPU performance, and streamlined Kubernetes orchestration at scale.

Purpose-designed for AI engineering teams, enterprises, and research institutions, Civo delivers direct access to cutting-edge NVIDIA GPU clusters, high-speed fabrics, and parallel storage systems required to train, fine-tune, and deploy foundation models efficiently.

We combine high-density infrastructure with predictable pricing and maximum compute throughput, empowering organizations to scale AI workloads without the complexity or cost bloat of traditional hyperscalers.

About the Role

As a Senior Solution Engineer – GPU & AI Infrastructure, you will serve as the primary technical architect for Civo’s large-scale AI and high-performance computing (HPC) customer initiatives.

You will be responsible for designing state-of-the-art NVIDIA GPU clusters tailored for training and inferencing massive foundation models.

In this role, you will bridge the gap between customer business objectives and ultra-high-performance hardware execution.

You will lead technical engagements, translate complex AI workload requirements into production-ready High-Level Designs (HLD), Low-Level Designs (LLD), and detailed Bills of Materials (BOM).

Your expertise will span bare-metal and Kubernetes-based orchestrations across cutting-edge NVIDIA Blackwell architectures (e. g., B300 and GB300NVL) using ultra-low-latency Infini Band and high-speed Ro CE networking fabrics.

Responsibilities

  • Solution Design & Architecture
  • System Design Documents: Author comprehensive High-Level Design (HLD) and Low-Level Design (LLD) documentation for enterprise-scale GPU supercomputing clusters.
  • Bill of Materials (BOM): Generate detailed BOMs covering compute nodes, NVLink switches, network fabrics, transceivers/cabling, liquid/air cooling requirements, power distribution, and high-performance storage.
  • GPU Cluster

Topology: Architect scale-up (NVLink/NVSwitch) and scale-out network topologies (Fat-Tree, Rail-Optimized) for NVIDIA Blackwell platforms, specifically B300 and GB300NVL rack-scale architectures.

  • Fabric & Networking

Engineering: Design high-throughput, low-latency networking architectures utilizing both Infini Band (e. g., NDR/X800) and Ro CE / Ro CEv2 (e. g., NVIDIA Spectrum-X / Spectrum-4) with lossless Ethernet mechanisms (PFC, ECN, Adaptive Routing).

  • Multi-Tenant & Deployment

Models: Deliver tailored architectures for both Bare-Metal (Slurm, Open MPI, bare-metal provisioning) and Cloud-Native / Kubernetes environments (NVIDIA GPU Operator, Network Operator, Run: ai, Kube Flow).

  • Storage Integration: Architect high-bandwidth parallel storage solutions utilizing GPUDirect Storage (GDS) and enterprise AI file systems (e. g., VAST Data).
  • Technical Sales Support & Customer Engagement
  • Partner with Civo’s sales and commercial teams as the technical lead for high-value AI infrastructure opportunities.
  • Engage directly with customer CTOs, Chief AI Officers, infrastructure leads, and ML engineers to evaluate technical requirements, compute sizing, and fabric choices.
  • Lead deep-dive architectural workshops and technical presentations on Civo's bare-metal GPU and managed Kubernetes offerings.
  • Produce precise technical proposals and lead responses to complex RFPs/RFIs regarding AI infrastructure.
  • Proof-of-Concept (Po C) & Benchmarking
  • Architect and oversee Proof-of-Concept (Po C) deployments to validate real-world performance for customer workloads.
  • Benchmark cluster performance using industry-standard tools (NCCL tests, GPUDirect RDMA latency/bandwidth, MLPerf, Megatron-LM benchmarks).
  • Address network congestion, fabric routing, and thermal/power optimization during validation phases.
  • Product & Ecosystem Collaboration
  • Serve as the bridge between enterprise AI clients, hardware vendors (NVIDIA, network OEMs), and Civo’s internal platform engineering team.
  • Provide continuous feedback to product teams on market trends, hardware platform demands, and feature requirements for AI/GPU orchestration.

Key Results/Objectives

  • Technical Wins: Achieve high technical win rates on large-scale AI/GPU cluster sales opportunities.
  • Design Excellence: Successfully deliver complete, peer-reviewed HLDs, LLDs, and BOMs within target deal timelines.
  • Customer Satisfaction: Achieve successful Po C completion and sign-off for enterprise clients scaling AI workloads on Civo infrastructure.

Requirements

Experience & Core Qualifications

  • 5+ years in a Solution Architecture, Systems Engineering, or Technical Pre-Sales role focused on high-performance cloud, HPC, or AI infrastructure.
  • Bachelor’s degree in Computer Science, Electrical Engineering, Systems Engineering, or equivalent practical experience.
  • Technical Expertise
  • NVIDIA GPU Architecture: Deep hands-on knowledge of NVIDIA HGX/DGX platforms, NVLink/NVSwitch fabrics, and Blackwell architectures (B300, GB300NVL, GB200 NVL72/NVL36).
  • High-Speed Networking: Expert-level knowledge of cluster fabric topologies:
  • Infini

Band: Quantum-2 / Quantum-X800, Subnet Management, Adaptive Routing.

  • Ro CE / Ro CEv2: Spectrum-X / Spectrum-4 Ethernet switches, PFC, ECN, Ro CE configuration, and optimization.
  • GPU Direct Technologies: GPUDirect RDMA (GDR) and GPUDirect Storage (GDS).
  • Orchestration & Platforms: Proficiency in deploying and optimizing GPU workloads on:
  • Kubernetes: Container networking (CNI), NVIDIA GPU Operator, RDMA Shared Device Plugin, MPI Operator.
  • Bare-Metal: Slurm, Ansible, Terraform, Py Torch/NCCL environment tuning.
  • Documentation Skills: Demonstrated experience creating enterprise-grade HLDs, LLDs, network rack diagrams, and itemized BOMs.
  • Power & Thermal

Awareness: Familiarity with high-density datacenter environments, liquid cooling technologies (Direct-to-Chip, CDU/liquid loop setups), and power delivery constraints for 100k W+ per rack deployments.

  • Soft Skills
  • Strong technical leadership and presentation skills, with the ability to articulate complex network and hardware tradeoffs to executive stakeholders.
  • Problem-solving mindset capable of diagnosing complex hardware-software interaction bottlenecks in distributed training/inference setups.
  • Location
  • Must be UK based.

Nice to Have

  • NVIDIA Certified Professional: AI Infrastructure (NCP-AII).
  • NVIDIA Certified Professional: AI Networking (NCP-AIN).
  • NVIDIA Certified Professional: Infini Band (NCP-IB).
  • NVIDIA Certified Associate / Professional: AI Workload Deployment & Cloud Native.

Why Join Civo?

  • Competitive compensation and benefits package.
  • 4-day week company (unless attending an event).
  • Uncapped holiday.
  • Remote work environment with flexibility and autonomy.
  • Collaborative and inclusive culture that values diversity and creativity.
  • Opportunity to work with a dynamic and innovative team in the fast-growing cloud industry.
  • #J-18808-Ljbffr

Senior Solution Engineer – GPU & AI Infrastructure employer: Civo Ltd

Civo is an exceptional employer, offering a competitive compensation package and a unique 4-day work week, fostering a healthy work-life balance. With a collaborative and inclusive culture that values diversity and creativity, employees have the opportunity to work with cutting-edge technology in the fast-growing cloud industry while enjoying flexibility and autonomy in a remote work environment. Civo also prioritises employee growth, providing avenues for professional development and engagement with a dynamic team dedicated to innovation in AI and HPC.

Civo Ltd

Contact Details:

Civo Ltd Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Senior Solution Engineer – GPU & AI Infrastructure

Join Local Tech Meetups

Get out there and mingle with fellow developers by joining local tech meetups. It’s a fantastic way to meet people who might be working at Civo Ltd or know someone who does. Plus, you can pick up some trendy tech skills and trends while you're at it!

Contribute to Open Source Projects

Show off your coding chops by jumping into open-source projects. Not only does this give you practical experience, but it also gets you noticed in the dev community. You'll create a killer portfolio that speaks volumes about your skills to Civo Ltd.

Tap into Online Developer Communities

Don’t underestimate the power of online developer communities like GitHub, Stack Overflow, and even Reddit. Participate in discussions, share your projects, and build your visibility. We can often find opportunities through these channels that can lead to a full-time gig at companies like Civo Ltd.

Explore Job Boards Specifically for Tech Roles

Keep your eyes peeled on job boards that focus on tech roles. Sites like TechCareers or Stack Overflow Jobs can often have listings for companies like Civo Ltd that might not show up on broader job sites. Make it a habit to check these regularly, and don’t hesitate to apply directly through our website!

We think you need these skills to ace Senior Solution Engineer – GPU & AI Infrastructure

NVIDIA GPU Architecture
High-Speed Networking
InfiniBand
RoCE / RoCEv2
GPUDirect Technologies
Kubernetes
Bare-Metal Provisioning

Some tips for your application 🫡

Show off your coding skills:When applying for a software engineering role, it's super important to showcase your coding skills. Make sure your CV includes your tech stack, any relevant programming languages you’re comfortable with, and examples of projects you've worked on. If you have a GitHub profile, link it up! We love to see code in action.

Tailor your portfolio:For a full-time role, we’d expect to see some solid examples of your work in your portfolio. Make sure to include at least two or three projects that highlight your problem-solving skills and your ability to work with different technologies. Focus on the projects that are most relevant to the position at Civo Ltd.

Craft a killer cover letter:Your cover letter is your chance to stand out—make it personal! Explain why you want to work at Civo Ltd and how your skills align with the role. Show us your passion for software development. We dig enthusiastic candidates who understand the value of collaboration and continuous learning!

Be clear and concise:When it comes to writing your CV and cover letter, clarity is key. Avoid jargon that could confuse us and stick to simple, direct language. Highlight your achievements with quantifiable results where possible, and keep everything easy to read. A well-organised application goes a long way!

How to prepare for a job interview at Civo Ltd

Brush Up on Your Coding Skills

For a full-time software engineering role, it's crucial that we stay sharp with our coding abilities. Expect technical questions that might involve solving problems on the spot or discussing algorithms. Practise on platforms like LeetCode or HackerRank to get comfortable with the types of questions that often come up.

Know Your Tools and Frameworks

Make sure we’re well-acquainted with the tools and technologies listed in the job description. Familiarise ourselves with any specific frameworks or programming languages mentioned. If Civo Ltd uses React or Node.js, for instance, be ready to discuss how we’ve used them in previous projects or coursework.

Showcase Your Projects

Bring along a portfolio that highlights our best work. This could be code samples, GitHub repositories, or any side projects we’ve built. Make sure we can talk through our thought process for each project, especially the challenges we faced and how we solved them—this shows our problem-solving skills in action.

Prepare for Behavioural Questions

While technical skills are key, full-time positions also require cultural fit. Be ready to discuss our previous experiences and how we handle teamwork, conflict, and deadlines. Brush up on the STAR method—Situation, Task, Action, Result—to clearly articulate our past experiences when discussing how we've contributed to a team.