Site Reliability Engineer

Site Reliability Engineer

Full-Time 36000 - 60000 £ / year (est.) No working from home possible
Duffel

At a Glance

  • Tasks: Ensure the reliability and performance of our innovative travel distribution infrastructure.
  • Company: Join Duffel, a forward-thinking tech company revolutionising the travel industry.
  • Benefits: Competitive salary, equity in the company, and a supportive growth-focused environment.
  • Other info: Diverse team culture with a commitment to equality and personal development.
  • Why this job: Make a real impact on global travel technology while developing your skills.
  • Qualifications: Experience in systems engineering, cloud platforms, and a passion for collaboration.

The predicted salary is between 36000 - 60000 £ per year.

Duffel is building tools to simplify travel distribution, search and booking, delivering one common and seamless API for hundreds of airlines. We aim to redesign the infrastructure underpinning the travel industry and improve the reliability and developer experience as we scale globally.

As an SRE, you’ll be part of a small engineering team responsible for the reliability, performance, and resilience of our infrastructure and applications. You will work closely with engineering teams to understand their needs and help meet the demands of our product as we scale globally.

What we’re looking for:

  • An infrastructure and systems engineering generalist who is comfortable diving deep into the weeds on different issues.
  • A configuration issue between Google’s Load Balancer and the HTTP server in our main Elixir application causing HTTP 5XX responses to customers.
  • Debugging an issue in our OpenTelemetry pipelines causing us to silently drop spans.
  • An enthusiasm for both software development and systems engineering.
  • A high bar for code and configuration quality and readability.
  • A good understanding of current observability and reliability practices.
  • Experienced and comfortable in running incident response.
  • Big picture thinking - ability to make trade-offs on technical work streams against business impact.
  • Excellent communication skills; the ability to articulate what you’re working on and why to the team in a clear and structured way.
  • Thrives in a collaborative environment, open to feedback and new ideas.
  • Experience with Google Cloud Platform (GCP) products such as GKE, Cloud SQL for PostgreSQL, BigQuery, Memorystore (Redis), etc.
  • Experience managing infrastructure and security for a PCI Cardholder Data Environment using Google Cloud Platform services and tooling.
  • Experience with Infrastructure as Code (Terraform).
  • Experience with a GitOps approach to Kubernetes using ArgoCD and Helm.
  • Experience with high-availability metrics collection systems (Grafana, Thanos, Prometheus) and transitioning to OpenTelemetry and Honeycomb for traces and metrics.
  • Experience with data pipelines (Pub/Sub, Airbyte, dbt).

Note: If your experience doesn’t exactly align with this stack, transferable skills will be considered. This description provides a sense of what you’ll be working with if you join the team.

What you can expect from us:

We’re dedicated to your personal growth. Our environment is supportive, and we value ideas, concerns and questions. Everyone who joins Duffel owns a share of the company and takes pride in their work.

Equality and recruitment:

We are an equal opportunities employer. We believe the key to our success is a diverse team; recruitment decisions are based on experience and skills. We welcome applications from everyone regardless of age, sex, disability, sexual orientation, race, religion or belief.

Note to recruitment agencies:

Duffel does not accept speculative CVs from external parties. Unsolicited CVs will be treated as the property of Duffel; attached terms and conditions are null and void.

Roles and locations:

Location: London, England, United Kingdom. This role is typically full-time and falls under the Seniority level: Mid-Senior level.

Site Reliability Engineer employer: Duffel

At Duffel, we are committed to making travel effortless and providing a dynamic work environment that fosters growth and innovation. Our London-based team enjoys a hybrid work model, generous benefits including a travel allowance and flexible bank holidays, and opportunities for professional development through study days and volunteering. Join us to be part of a diverse team that values your skills and creativity while working at the forefront of the travel industry.

Duffel

Contact Details:

Duffel Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Site Reliability Engineer

Tip Number 1

Network like a pro! Reach out to current or former employees at Duffel on LinkedIn. A friendly chat can give you insider info and maybe even a referral, which can really boost your chances.

Tip Number 2

Prepare for the technical interview by brushing up on your SRE skills. Dive into topics like incident response and observability practices. We want to see your problem-solving skills in action!

Tip Number 3

Show off your passion for both software development and systems engineering. During interviews, share examples of how you've tackled complex issues and improved system reliability. Let your enthusiasm shine through!

Tip Number 4

Don’t forget to apply through our website! It’s the best way to ensure your application gets seen by the right people. Plus, it shows you’re genuinely interested in joining the Duffel team.

We think you need these skills to ace Site Reliability Engineer

Infrastructure and Systems Engineering
Google Cloud Platform (GCP)
Elixir
OpenTelemetry
Incident Response
Code Quality
Observability Practices

Some tips for your application 🫡

Tailor Your CV:Make sure your CV reflects the skills and experiences that align with the Site Reliability Engineer role. Highlight your experience with Google Cloud Platform, Infrastructure as Code, and any relevant debugging or incident response work you've done.

Craft a Compelling Cover Letter:Use your cover letter to tell us why you're passionate about SRE and how your background makes you a great fit for Duffel. Be sure to mention specific projects or experiences that showcase your problem-solving skills and enthusiasm for both software development and systems engineering.

Showcase Your Communication Skills:Since excellent communication is key for this role, make sure your application materials are clear and well-structured. We want to see how you articulate your thoughts and ideas, so don’t shy away from explaining your past projects and their impact.

Apply Through Our Website:We encourage you to apply directly through our website. This not only streamlines the process but also shows us that you're genuinely interested in joining our team at Duffel. Plus, it’s the best way to ensure your application gets into the right hands!

How to prepare for a job interview at Duffel

Know Your Tech Stack

Make sure you’re familiar with the technologies mentioned in the job description, especially Google Cloud Platform and Infrastructure as Code tools like Terraform. Brush up on your knowledge of observability practices and be ready to discuss how you've used these tools in past projects.

Prepare for Problem-Solving Questions

Expect to tackle real-world scenarios during the interview. Think about how you would debug issues like configuration problems or incident responses. Practise articulating your thought process clearly, as communication is key in a collaborative environment.

Show Your Enthusiasm for Collaboration

Duffel values teamwork, so be prepared to share examples of how you’ve worked effectively with others. Highlight any experiences where you’ve received or given feedback, and how that contributed to project success. This will demonstrate your fit within their culture.

Think Big Picture

Be ready to discuss how technical decisions impact business outcomes. Prepare to explain trade-offs you’ve made in previous roles and how they aligned with broader company goals. This shows you understand the importance of balancing technical work with business needs.