Senior Backend Engineer, LLM Inference Platform

Senior Backend Engineer, LLM Inference Platform

Full-Time 80000 - 100000 £ / year (est.) No working from home possible

At a Glance

  • Tasks: Design and build backend services for scalable LLM inference in a dynamic team.
  • Company: Join J.P. Morgan, a leading financial institution with a focus on innovation.
  • Benefits: Competitive salary, health benefits, and opportunities for professional growth.
  • Other info: Work in an agile environment with excellent career advancement potential.
  • Why this job: Make an impact by optimising LLM platforms and enhancing performance.
  • Qualifications: Experience in software engineering and a passion for backend development.

The predicted salary is between 80000 - 100000 £ per year.

J.P. Morgan is seeking a Software Engineer III to join the Firmwide LLM Serving Platform. You will design, build, and operate backend services for scalable LLM inference, contributing to a production platform that reduces latency, increases throughput, and maximizes GPU utilization.

In this role based in Greater London, you will work in an agile team, learn about model architectures, observability, CI/CD, and secure coding practices, and collaborate across engineering, product, and operations.

Senior Backend Engineer, LLM Inference Platform employer: 慨正橡扯

At 慨正橡扯, we pride ourselves on being an exceptional employer that champions innovation and collaboration in the field of Behavioral Economics and Retirement Research. Our hybrid working model not only offers flexibility but also nurtures a vibrant work culture where employees are encouraged to grow and develop their skills through meaningful projects and leadership opportunities. Join us in Europe, where your expertise will directly contribute to enhancing investor outcomes and shaping impactful business strategies.

Contact Details:

慨正橡扯 Recruitment Team

We think you need these skills to ace Senior Backend Engineer, LLM Inference Platform

Backend Development
LLM Inference
Scalability
Latency Reduction
Throughput Optimization
GPU Utilization
Agile Methodologies