At a Glance
- Tasks: Train and fine-tune cutting-edge video models to create amazing AI-generated content.
- Company: Join VEED, a leading generative AI company revolutionising video creation for social media.
- Benefits: Enjoy unlimited holidays, flexible subsidies, and mental health support.
- Other info: Work remotely or from our vibrant hubs in London, Barcelona, or Amsterdam.
- Why this job: Be part of a dynamic team shaping the future of video with innovative technology.
- Qualifications: Experience with ML models and a passion for impactful AI solutions.
The predicted salary is between 76000 - 119400 £ per year.
About The Company
At VEED, we're a generative AI company building the AI video creation platform made for social. It's where millions of marketers imagine and generate branded videos, in minutes. We're powered by frontier models and VEED Fabric, our own image-to-video model. Our APIs also run AI video inside leading creative products. We're 150 people, $50M ARR, backed by Sequoia. We are hiring. Come help us make AI videos worth posting.
Where And How We Work
We’re a distributed team with hubs in London, Barcelona, and Amsterdam. Our ML team is mainly in London, so we’re open to candidates already based there, willing to relocate, or working remotely within 4–5 hours of London time (GMT).
About The Team
You’ll be joining a team of experienced ML engineers and researchers (ex-Spiritme, ex-Pipio) working on scaling our generative video models. The team’s goal is to turn breakthroughs in research into fast, reliable, and accessible systems that power VEED’s future.
About The Role
We’ve just released Fabric — a DiT-based speech-to-video model that turns spoken words into coherent, human‑centric video. We’re now building the next version and a suite of video generation & editing models focused on controllability, quality, and speed. You will work across the full model lifecycle for Fabric 2 and related models — from exploring new research and approaches to training and fine‑tuning models. Your work will turn ideas from research into features creators use every day.
What You'll Do
- Training and fine‑tuning models to create best‑in‑class generation and editing people‑centric video models
- Running experiments to boost visual quality, lip sync, and controllability
- Defining clear metrics and evaluation loops and using them to guide decisions
- Work closely with infra and performance engineers and collaborate with the product team
Our Stack
Model/Training: PyTorch, diffusion/DiT, diffusers, transformers, torch.compile, torch.distributed, torchrun, Accelerate, PEFT/LoRA. Data: Python, Hugging Face datasets, WebDataset‑style storage, ffmpeg, large‑scale curation & filtering. Serving: Optimised PyTorch/TensorRT runtimes in GPU‑backed Python services; experiment tracking, A/B testing, eval pipelines. Hardware: Access to latest‑gen NVIDIA GPUs H100 and B200 for training and inference.
About You
- Experience training diffusion or transformer models for video, vision, or audio with results in production or papers
- You care about shipping useful ML, not just benchmarks.
- You enjoy owning problems and collaborating across teams.
- Strong data and experimentation skills and confidence designing practical metrics
- Comfortable writing clean Python and reviewing research with a pragmatic eye
- Clear communicator who can explain trade‑offs and make sound decisions
- Motivated by impact and by making generative video better, faster, and more accessible
What We Offer
- Monthly subsidy programme: Different people have different needs and therefore value different benefits. Providing this as a subsidy allows you to have the greatest flexibility to apply when you value most - whether that be to offset the cost of office furniture, childcare, gym membership, etc.
- Unlimited paid holidays: We value that you get more time with your family and friends.
- Home office set‑up: We have an IT Equipment programme to make sure your home office is adequately set up with IT equipment including a laptop, monitors, headsets/earbuds, keyboards and more!
- Mental health benefit: We’ve partnered with Spill to provide all our employees with confidential mental health support.
This role is open at IC3 or IC4 level depending on experience. The salary range across both levels is £76,000 - £119,400 gross per year. The level and offer are determined by the scope of the role and your experience and skills assessed during the process. We do not ask about current or previous salary at any stage. We think what matters is people. After all, a company is just a group of people. We don’t care about where you’re from, what school you went to or where you worked before. If you’ve done exceptional work, we want to hear from you. Join us on our mission to make creative storytelling with video simple and accessible for everyone.
Country Hiring Guidelines
At VEED we are hybrid, enabling teams and individuals to design their day and integrate work and life. We are currently hiring in 3 core hubs: London, Amsterdam and Barcelona and 2 additional hubs for sales and support roles only (USA for sales and the Philippines for support). Please refer to the individual job posts for more details.
By submitting this application, I agree that my personal data will be collected, processed, and retained by the company solely for the purposes of managing and assessing my candidacy.
ML Engineer employer: VEED
At VEED, we pride ourselves on being an innovative employer that champions creativity and collaboration within a dynamic work culture. Our London hub is a vibrant space where ML engineers can thrive, supported by flexible benefits like a monthly subsidy programme, unlimited paid holidays, and comprehensive mental health support. Join us to be part of a forward-thinking team dedicated to making generative video accessible and impactful for everyone.