At a Glance
- Tasks: Join us in deploying groundbreaking AI technology and support customer integration.
- Company: Exciting UK startup revolutionising AI with innovative optical computing.
- Benefits: Competitive salary, share options, health insurance, and generous holiday allowance.
- Other info: Dynamic team culture with excellent growth opportunities and inclusive community events.
- Why this job: Be at the forefront of AI innovation and make a real impact.
- Qualifications: Experience in software engineering, Python skills, and familiarity with ML frameworks.
The predicted salary is between 80000 - 100000 £ per year.
The Opportunity
Lumai is redefining how the world computes. We are an ambitious, venture-backed UK startup pioneering a breakthrough AI accelerator for data centers which uses 3D optical compute. Our radical technology uses light to perform computation at orders of magnitude faster speeds and at far greater scales than ever before, all whilst consuming far less energy than traditional approaches. Lumai is unlocking performance and efficiency gains that could transform the economics of AI and compute infrastructure and reshape how intelligence scales globally. If you are passionate about bringing groundbreaking technology to market, and want to be part of a team pushing the boundaries of what is physically possible, Lumai is where you can make it happen.
About Lumai
Founded in 2022, Lumai is a University of Oxford spinout using optical processing to accelerate large language models (LLMs) and other transformer-based AI systems. The team combines expertise in optical computing, machine learning, and physics. Lumai has already secured over $15 million in investment from leading deep-tech investors like Constructor Capital, IP Group, PhotonVentures and government grants, and is scaling rapidly to deploy the fastest optical compute currently available globally.
The Role
We are bringing the world's first optical AI compute platform to market. As we move from development into field deployment, we are looking for a Software Inference Deployment Engineer to own the software-side integration and customer support of Lumai Iris servers in third-party data centre environments. You will begin by working alongside our software and engineering teams - helping integrate the Iris software stack, supporting model onboarding through the toolchain, and getting hands‑on with the disaggregated prefill/decode runtime. This is intentional: the best way to develop deep expertise in a novel platform is to build with it. As deployments go live, you will take ownership in the field - supporting customer integration into their inference stacks, troubleshooting software issues, and acting as a primary technical contact for customer ML and infrastructure engineering teams. This is an opportunity to work at the cutting edge of efficient AI inference - deploying a genuinely novel compute platform into production for the first time, and playing a central role in how it reaches the world.
What You’ll Do
- Work alongside Lumai's software and engineering teams to integrate, test, and harden the Iris software stack ahead of deployment
- Support model onboarding through the Iris toolchain - loading, conversion, and framework integration
- Develop hands‑on familiarity with the disaggregated prefill/decode runtime, including how Iris servers operate alongside decode processors
- Support customer integration of Lumai Iris into their own frameworks
- Own software‑side troubleshooting in the field, acting as the first line of response post‑deployment
- Train and enable customer ML and infrastructure engineering teams on the Iris software platform
- Feed field findings, integration issues, and customer feedback back into product and engineering
What We're Looking For
Must‑Have
- Hands‑on software engineering experience in AI infrastructure, inference serving, accelerator integration, or comparable deep‑tech hardware‑software environments
- Strong Python skills and familiarity with major ML frameworks (PyTorch in particular)
- Practical experience with model deployment workflows - loading, format conversion, quantisation, or framework integration
- Comfortable working with inference serving stacks (for example vLLM, TensorRT‑LLM, or similar)
- Familiarity with Linux, containerisation (Docker), and cluster environments
- Comfortable in a customer‑facing role, able to communicate clearly with ML and infrastructure engineering teams
- Comfortable working in a fast‑moving, early‑stage environment where the product and the deployment approach are both still being developed
Strong Preference For
- Experience integrating accelerator hardware (GPUs, FPGAs, ASICs, NPUs, or novel architectures) into customer inference workflows
- Familiarity with the NVIDIA inference stack - CUDA, TensorRT, Triton
- Exposure to disaggregated inference architectures, prefill/decode separation, or KV cache management
Compensation & Benefits
- Highly Competitive Salary: We are not saying our salary is a blank check, but let's just say it won't be a source of your stress
- Share Option Scheme: We are all in this together! We believe in shared success while we build the Lumai of tomorrow
- Pension Scheme: Plan for retirement with AVIVA
- Private Health Insurance: We firmly believe that you come first, and a happy you is a healthy you! Look after yourself and your loved ones with AXA
- Cycle to Work: Spread the cost of a bike, a bike and accessories or just accessories and save on tax
- L&D Allowance: Stay at the forefront of your field with a £500 annual development budget
- Subsidised On‑site Lunches: Enjoy on‑site healthy meals at half the price, as Lumai covers 50% of the cost
- Holidays: Enjoy some deserved "me time" with 25 days paid holiday (plus bank holidays) per year
- Socials: Be part of an inclusive community enjoying occasional all‑company off‑sites, lunches and socials
Interview Process
Our process is four stages. An initial conversation with our HR team to understand what you want from the role and what we want from it. Two technical sessions with our Product and Leadership team. Finally, an HR‑team session covering scope, terms, and any final questions. We aim to move fast on candidates we are excited about; expect roughly three to four weeks end to end.
Lumai is an equal opportunity employer. We make hiring decisions on merit, scope‑fit, and the strength of the working relationship we expect to build with each hire. Applications welcome from candidates of any background. If you are not sure whether you are a fit, send a note anyway.
Software Inference Deployment Engineer employer: Lumai Limited
At Lumai, we are not just redefining computation; we are creating a vibrant and inclusive work culture that fosters innovation and collaboration. As a rapidly growing UK startup, we offer competitive salaries, share options, and a generous learning and development allowance, ensuring our employees have the resources to thrive. Join us in Oxford, where you will have the unique opportunity to shape groundbreaking technology while enjoying a supportive environment that values your contributions and promotes personal growth.
StudySmarter Expert Advice🤫
We think this is how you could land Software Inference Deployment Engineer
✨Join Local Tech Meetups
Get out there and mingle with fellow developers by joining local tech meetups. It’s a fantastic way to meet people who might be working at Lumai Limited or know someone who does. Plus, you can pick up some trendy tech skills and trends while you're at it!
✨Contribute to Open Source Projects
Show off your coding chops by jumping into open-source projects. Not only does this give you practical experience, but it also gets you noticed in the dev community. You'll create a killer portfolio that speaks volumes about your skills to Lumai Limited.
✨Tap into Online Developer Communities
Don’t underestimate the power of online developer communities like GitHub, Stack Overflow, and even Reddit. Participate in discussions, share your projects, and build your visibility. We can often find opportunities through these channels that can lead to a full-time gig at companies like Lumai Limited.
✨Explore Job Boards Specifically for Tech Roles
Keep your eyes peeled on job boards that focus on tech roles. Sites like TechCareers or Stack Overflow Jobs can often have listings for companies like Lumai Limited that might not show up on broader job sites. Make it a habit to check these regularly, and don’t hesitate to apply directly through our website!
We think you need these skills to ace Software Inference Deployment Engineer
Some tips for your application 🫡
Show off your coding skills:When applying for a software engineering role, it's super important to showcase your coding skills. Make sure your CV includes your tech stack, any relevant programming languages you’re comfortable with, and examples of projects you've worked on. If you have a GitHub profile, link it up! We love to see code in action.
Tailor your portfolio:For a full-time role, we’d expect to see some solid examples of your work in your portfolio. Make sure to include at least two or three projects that highlight your problem-solving skills and your ability to work with different technologies. Focus on the projects that are most relevant to the position at Lumai Limited.
Craft a killer cover letter:Your cover letter is your chance to stand out—make it personal! Explain why you want to work at Lumai Limited and how your skills align with the role. Show us your passion for software development. We dig enthusiastic candidates who understand the value of collaboration and continuous learning!
Be clear and concise:When it comes to writing your CV and cover letter, clarity is key. Avoid jargon that could confuse us and stick to simple, direct language. Highlight your achievements with quantifiable results where possible, and keep everything easy to read. A well-organised application goes a long way!
How to prepare for a job interview at Lumai Limited
✨Brush Up on Your Coding Skills
For a full-time software engineering role, it's crucial that we stay sharp with our coding abilities. Expect technical questions that might involve solving problems on the spot or discussing algorithms. Practise on platforms like LeetCode or HackerRank to get comfortable with the types of questions that often come up.
✨Know Your Tools and Frameworks
Make sure we’re well-acquainted with the tools and technologies listed in the job description. Familiarise ourselves with any specific frameworks or programming languages mentioned. If Lumai Limited uses React or Node.js, for instance, be ready to discuss how we’ve used them in previous projects or coursework.
✨Showcase Your Projects
Bring along a portfolio that highlights our best work. This could be code samples, GitHub repositories, or any side projects we’ve built. Make sure we can talk through our thought process for each project, especially the challenges we faced and how we solved them—this shows our problem-solving skills in action.
✨Prepare for Behavioural Questions
While technical skills are key, full-time positions also require cultural fit. Be ready to discuss our previous experiences and how we handle teamwork, conflict, and deadlines. Brush up on the STAR method—Situation, Task, Action, Result—to clearly articulate our past experiences when discussing how we've contributed to a team.