GenAI Infra Engineer: Real-Time GPU Serving & Fine-Tuning

GenAI Infra Engineer: Real-Time GPU Serving & Fine-Tuning

Full-Time 80000 - 100000 Β£ / year (est.) No working from home possible
United States Digital Space LLC

At a Glance

  • Tasks: Build and optimise GenAI infrastructure for real-time GPU serving and model fine-tuning.
  • Company: Join a dynamic team at the forefront of AI technology.
  • Benefits: Competitive pay, flexible hours, and opportunities for skill development.
  • Other info: Collaborative environment with potential for rapid career advancement.
  • Why this job: Be part of groundbreaking projects that shape the future of AI.
  • Qualifications: Experience in software engineering and machine learning infrastructure.

The predicted salary is between 80000 - 100000 Β£ per year.

We are hiring a Software Engineer, Machine Learning Infrastructure to join a small, high-leverage team building production Gen AI infrastructure at scale.

The role focuses on real-time GPU serving, high-throughput batch inference, and model fine-tuning for open-weight platforms.

You will collaborate across model serving, inference engines, training pipelines, and observability, pushing cost/performance frontiers while meeting latency and reliability targets.

#J-18808-Ljbffr

GenAI Infra Engineer: Real-Time GPU Serving & Fine-Tuning employer: United States Digital Space LLC

United States Digital Space LLC is an exceptional employer, offering a dynamic work culture that prioritises innovation and collaboration in the heart of Greater London. With a strong focus on employee well-being and flexible work options, we provide ample opportunities for professional growth and development, making it an ideal environment for those looking to make a meaningful impact in the field of AI-enabled SaaS engineering.

United States Digital Space LLC

Contact Details:

United States Digital Space LLC Recruitment Team

We think you need these skills to ace GenAI Infra Engineer: Real-Time GPU Serving & Fine-Tuning

Real-Time GPU Serving
High-Throughput Batch Inference
Model Fine-Tuning
Machine Learning Infrastructure
Collaboration Skills
Cost/Performance Optimisation
Latency Management