At a Glance
- Tasks: Lead a DevOps team to build and optimise infrastructure for AI deployment.
- Company: Join a remote-first tech collective focused on innovation and collaboration.
- Benefits: Flexible hours, competitive salary, and opportunities for personal growth.
- Other info: Enjoy a dynamic work environment with regular downtime to recharge.
- Why this job: Shape the future of AI infrastructure while working with cutting-edge technologies.
- Qualifications: Experience in DevOps, automation, and system reliability is essential.
The predicted salary is between 60000 - 80000 £ per year.
Runware’s infrastructure is the foundation that enables our teams to deliver AI to the world. As DevOps Team Lead, you’ll turn complex, hardware-driven systems into streamlined, developer-friendly platforms. You’ll define how we automate deployments, orchestrate GPUs at scale, and observe workloads in real time. You’ll build systems that detect and recover from issues before users notice, and work closely with engineering and product teams to make shipping faster, safer, and more predictable. You’ll shape the foundation that lets teams move fast with confidence, building infrastructure that is dependable, observable, and designed to scale.
Is this role a fit for you? You thrive at the intersection of infrastructure and innovation. You enjoy unravelling complex systems, tuning performance, and engineering reliability into everything you build. You lead through clarity and example, not process, and elevate the teams around you by simplifying the hard things. You take pride in building systems that are resilient by design and empowering the engineers who depend on them. You understand that reliability is never accidental; it is built through intent, consistency, and a culture that values doing things right.
What this role will entail:
- Providing technical and people leadership to a small DevOps team
- Lead the design and operation of Runware’s infrastructure and orchestration systems
- Build automation and tooling to streamline model deployments, scaling, and hardware utilisation across distributed nodes
- Drive observability, alerting, and reliability practices to detect and resolve issues quickly and proactively
- Collaborate with engineers to optimise throughput, latency, and platform performance at every layer of the stack
- Develop and maintain infrastructure as code and deployment automation to ensure consistency and reproducibility across environments
- Establish and continuously evolve incident management, post-mortems, and reliability reviews as core engineering practices
- Mentor and coach engineers to think operationally, designing systems that fail gracefully and scale predictably
- Champion forward-looking improvements to our orchestration layer, hardware management, and overall infrastructure efficiency
- Have experience operating production systems on bare metal or hybrid environments such as HPC or GPU clusters, optimised for performance and low latency
- Are comfortable writing automation and systems tooling in Python, Go, or similar languages
- Understand container runtimes like Docker and containerd, and have built or worked with orchestration systems beyond Kubernetes
- Are fluent in observability and debugging practices across distributed systems, using logs, metrics, traces, and profiling to drive insight and reliability
- Care deeply about reliability, efficiency, and engineering quality, and know how to embed those values into team culture and everyday practice
- Thrive in fast-moving, evolving environments where impact is measured by how much better systems and teams perform over time
We’re a remote-first collective, meeting in person twice a year to plan, brainstorm, celebrate wins, and enjoy some face-to-face time. We have core hours for cooperative working and calls, but outside of that your calendar is yours. Work the hours that let you perform at your peak while also building a healthy life. Our release cycles are fast and intense, but they’re followed by real downtime. After big pushes we expect the team to unplug, recharge, and come back ready.
Remote DevOps Team Lead in Plymouth employer: Runware
Runware is an exceptional employer for those looking to lead in the DevOps space, offering a remote-first work culture that prioritises flexibility and work-life balance. With generous paid time off, meaningful stock options, and opportunities for professional growth through mentorship, employees are empowered to thrive in a fast-paced environment while enjoying the benefits of collaborative retreats twice a year. Join us to shape innovative infrastructure solutions that drive AI delivery globally, all while working from anywhere in the UK.
StudySmarter Expert Advice🤫
We think this is how you could land Remote DevOps Team Lead in Plymouth
✨Tip Number 1
Network like a pro! Reach out to folks in the industry, join relevant online communities, and attend meetups. You never know who might have the inside scoop on job openings or can refer you directly.
✨Tip Number 2
Show off your skills! Create a portfolio or GitHub repository showcasing your projects and contributions. This is your chance to demonstrate your expertise in DevOps and make a lasting impression.
✨Tip Number 3
Prepare for interviews by practising common DevOps scenarios and technical questions. Mock interviews with friends or using online platforms can help you feel more confident and ready to tackle any challenge.
✨Tip Number 4
Don’t forget to apply through our website! It’s the best way to ensure your application gets noticed. Plus, we love seeing candidates who are genuinely interested in joining our remote-first team.
We think you need these skills to ace Remote DevOps Team Lead in Plymouth
Some tips for your application 🫡
Show Your Passion for Infrastructure:When you're writing your application, let your enthusiasm for infrastructure and innovation shine through. We want to see how you’ve tackled complex systems in the past and how you can bring that experience to our team.
Be Clear and Concise:Keep your application straightforward and to the point. We appreciate clarity, so make sure to highlight your relevant skills and experiences without unnecessary fluff. This helps us see your potential as a DevOps Team Lead right away!
Tailor Your Application:Make sure to customise your application for this specific role. Mention how your background aligns with our needs, especially around automation, observability, and reliability practices. We love seeing candidates who take the time to connect their experience with what we’re looking for.
Apply Through Our Website:Don’t forget to submit your application through our website! It’s the best way for us to keep track of your application and ensure it gets the attention it deserves. Plus, it shows you’re serious about joining our remote-first collective!
How to prepare for a job interview at Runware
✨Know Your Tech Inside Out
Make sure you’re well-versed in the technologies mentioned in the job description, like Python, Go, and container runtimes. Brush up on your experience with orchestration systems and be ready to discuss how you've optimised performance in past roles.
✨Showcase Your Leadership Style
As a DevOps Team Lead, your leadership approach is crucial. Prepare examples of how you've led teams through complex challenges, simplified processes, and fostered a culture of reliability and efficiency. Highlight your mentoring experiences and how you've empowered others.
✨Demonstrate Problem-Solving Skills
Be ready to tackle hypothetical scenarios during the interview. Think about how you would handle incidents or optimise systems under pressure. Use the STAR method (Situation, Task, Action, Result) to structure your responses and showcase your analytical thinking.
✨Align with Company Culture
Research Runware’s values and work culture. Be prepared to discuss how you thrive in fast-moving environments and your approach to maintaining a healthy work-life balance. Show that you understand the importance of downtime and team collaboration in a remote-first setting.