Site Reliability Engineer (SRE)

Site Reliability Engineer (SRE)

Full-Time 36000 - 60000 £ / year (est.) Home office (partial)
D

At a Glance

  • Tasks: Ensure the reliability and performance of our travel tech infrastructure while collaborating with engineering teams.
  • Company: Join Duffel, a pioneering company transforming the future of travel technology.
  • Benefits: Competitive salary, flexible working hours, and opportunities for professional growth.
  • Other info: Work with cutting-edge technologies like Google Cloud Platform and Kubernetes in a dynamic team.
  • Why this job: Be part of a mission to simplify travel and make a real impact in the industry.
  • Qualifications: Experience in systems engineering, strong coding skills, and excellent communication.

The predicted salary is between 36000 - 60000 £ per year.

Overview

Create the future of travel with us. Whether it’s to visit the people closest to us, starting an exciting adventure, or a career-defining business trip, travel is an essential part of our lives. Yet we’ve all experienced the aches and pains of getting to our destination. Today, more than 4 billion airline passengers rely on technology that hasn’t kept up with the expectations of the modern connected traveller. That’s why we’ve started to rebuild the infrastructure that underpins the travel industry. We’re on a mission to unravel travel — simplifying systems and building the tools that will make the future of travel effortless.

Engineering at Duffel

We’re building tools to simplify travel distribution, search and booking. What does this actually mean? It’s one common and seamless API. This brings huge technical challenges as we need to design and build a beautiful API before integrating to hundreds of airlines. Along with that we need to navigate through the differing needs and systems of each airline whilst building a fantastic developer experience to go with it. The tools used on the team include Elixir, Phoenix, Kubernetes and Google Cloud Platform.

Site Reliability Engineering at Duffel

As an SRE at Duffel, you’ll be part of a small team within engineering that is responsible for the reliability, performance, and resilience of our infrastructure and applications. You will be working closely with engineering teams to understand their needs and help meet the demands of our product as we scale globally.

What We’re Looking For

  • An infrastructure and systems engineering generalist who is comfortable diving deep into the weeds on different issues.
  • Some recent examples include:
    • A configuration issue between Google’s Load Balancer and the HTTP server in our main Elixir application causing HTTP 5XX responses to be returned to our customers.
    • Debugging an issue in our OpenTelemetry pipelines causing us to silently drop spans.
  • An enthusiasm for both software development and systems engineering.
  • A high bar for code and configuration quality and readability.
  • A good understanding of current observability and reliability practices.
  • Experienced and comfortable in running incident response.
  • Big picture thinking - you can make trade offs on technical work streams against business impact.
  • Fantastic communication skills. You’re able to articulate what you’re working on and why to the team in a clear and structured way.
  • You thrive in a collaborative environment. You believe in your own methods but keep an open mind, taking suggestions and feedback onboard as well.

Technologies

We run our infrastructure on Google Cloud Platform, so you’ll be helping to run a few of their products such as GKE, CloudSQL for PostgreSQL, BigQuery, Memorystore (Redis) and more. We manage the infrastructure and security for a segregated PCI Cardholder Data Environment, entirely managed with Google Cloud Platform services and tooling. We follow an Infrastructure as Code approach to managing our infrastructure, using Terraform. We follow a GitOps approach to managing our Kubernetes configuration, using ArgoCD and Helm. We manage a high-availability metrics collection system using Grafana, Thanos.

Site Reliability Engineer (SRE) employer: Duffel

At Duffel, we are committed to making travel effortless and providing a dynamic work environment that fosters growth and innovation. Our London-based team enjoys a hybrid work model, generous benefits including a travel allowance and flexible bank holidays, and opportunities for professional development through study days and volunteering. Join us to be part of a diverse team that values your skills and creativity while working at the forefront of the travel industry.

D

Contact Details:

Duffel Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Site Reliability Engineer (SRE)

Tip Number 1

Network like a pro! Reach out to folks in the industry, attend meetups, and connect with current SREs. You never know who might have the inside scoop on job openings or can refer you directly.

Tip Number 2

Show off your skills! Create a portfolio or GitHub repository showcasing your projects, especially those involving Elixir, Kubernetes, or Google Cloud Platform. This gives potential employers a taste of what you can do.

Tip Number 3

Prepare for technical interviews by brushing up on your incident response strategies and observability practices. Be ready to discuss real-world scenarios and how you tackled them—this is your chance to shine!

Tip Number 4

Don’t forget to apply through our website! It’s the best way to ensure your application gets seen by the right people. Plus, we love seeing candidates who are proactive about their job search.

We think you need these skills to ace Site Reliability Engineer (SRE)

Elixir
Phoenix
Kubernetes
Google Cloud Platform
Infrastructure as Code
Terraform
GitOps

Some tips for your application 🫡

Tailor Your CV:Make sure your CV reflects the skills and experiences that align with the Site Reliability Engineer role. Highlight your experience with technologies like Google Cloud Platform, Kubernetes, and any relevant incident response work. We want to see how you can contribute to our mission!

Craft a Compelling Cover Letter:Your cover letter is your chance to shine! Use it to explain why you're passionate about travel tech and how your background makes you a great fit for our team. Be sure to mention specific projects or experiences that showcase your problem-solving skills.

Showcase Your Communication Skills:As an SRE, you'll need to communicate complex ideas clearly. In your application, demonstrate your ability to articulate technical concepts in a straightforward way. This could be through examples of past collaborations or how you've explained technical issues to non-technical stakeholders.

Apply Through Our Website:We encourage you to apply directly through our website. It’s the best way for us to receive your application and ensures you’re considered for the role. Plus, it shows you’re keen on joining our team at Duffel!

How to prepare for a job interview at Duffel

Know Your Tech Stack

Familiarise yourself with the technologies mentioned in the job description, like Elixir, Kubernetes, and Google Cloud Platform. Be ready to discuss how you've used these tools in past projects or how you would approach challenges using them.

Showcase Problem-Solving Skills

Prepare to share specific examples of how you've tackled complex issues in infrastructure or systems engineering. Think about times when you resolved configuration problems or improved system reliability, and be ready to explain your thought process.

Communicate Clearly

Practice articulating your ideas and experiences in a clear and structured manner. Since communication is key in a collaborative environment, consider doing mock interviews with friends or colleagues to refine your delivery.

Emphasise Collaboration

Highlight your experience working in teams and how you value feedback. Discuss instances where you’ve successfully collaborated with others to achieve a common goal, as this aligns with the company’s emphasis on teamwork.