Site Reliability Engineer - APAC in London

Site Reliability Engineer - APAC in London

London Full-Time 60000 - 80000 £ / year (est.) No working from home possible
T

At a Glance

  • Tasks: Join us to optimise and automate our global Cloud platform while solving real-world problems.
  • Company: Tyk, a leading API Management platform with a mission to connect every system in the world.
  • Benefits: Unlimited paid holidays, remote work, employee share scheme, and generous parental leave.
  • Other info: Embrace a culture of authenticity, respect, and continuous improvement.
  • Why this job: Be part of a dynamic team that values innovation and flexibility in a supportive environment.
  • Qualifications: Experience in SRE, cloud technologies, and strong communication skills are essential.

The predicted salary is between 60000 - 80000 £ per year.

Who are Tyk, and what do we do? The Tyk API Management platform is helping to drive the connected world and power new products and services. We’re changing the way that organisations connect any number of their systems and services. Whether internal, external, public or highly encrypted systems, Tyk helps businesses drive value across various industries.

Founded in 2015 with offices in London - UK, London - Ontario, Atlanta and Singapore, we have many thousands of users of our B2B platform across the globe. Our Mission is to connect every system in the world by building an API Management platform.

We offer unlimited paid holidays and remote working from anywhere in the world for everyone. Tyk was founded on the principle of offering flexibility and autonomy to our employees, which lets them achieve their best results.

The role involves:

  • Lead hands-on maintenance and optimisation of our global Cloud platform within SL(A/I/O)s.
  • Collaborate to shape SRE strategy, then translate into actionable technical plans coordinated through SCRUM.
  • Identify reliability issues, drive root-cause analysis, and implement solutions alongside your squad.
  • Lead performance tuning and fault finding through analysis of OS and application metrics.
  • Design and implement automation for common operational tasks and cloud-operations workflows.
  • Develop proactive alerting, monitoring roadmap, and relevant dashboards; define and track KPIs.
  • Participate in on-call rotation, ensuring effective incident response and resolution within SLAs.
  • Conduct blame-free post-mortems, document findings, and maintain operational runbooks.
  • Drive multi-region and multi-cloud platform expansion with focus on scalability and automation.
  • Optimise infrastructure performance and cost efficiency without impacting service delivery.
  • Engage with commercial teams on growth plans and translate into technical SRE strategies.
  • Coordinate penetration testing through provider liaison, technical setup, and environment configuration.
  • Champion continuous improvement across processes, communication, and team practices.
  • Model excellence in software design and knowledge sharing.
  • Plan and execute software upgrades to enhance cloud services.

Experience required:

  • Experience in an SRE role.
  • Strong knowledge of cloud technologies and SLA / SLO / SLI management.
  • Excellent communication and leadership skills.
  • Ability to analyse and improve operational processes and performance metrics.
  • Experience in software design, automation, and root-cause analysis.
  • On-call support experience and customer-focused mindset.
  • Collaborative attitude with commercial and technical teams.
  • Launching and operating production Kubernetes clusters.
  • Designing and operating infrastructure on AWS and other providers.
  • Operating MongoDB (or other document database) clusters.
  • Operating Redis (or other key-value storage) clusters.
  • Administering Linux servers.
  • Operating Prometheus and Grafana.
  • Operating logging collection and analysis system.
  • Participating in the on-call rotation (4:00am – 16:00pm UTC).

Skills:

  • Kubernetes (administrator).
  • Go and/or Python (advanced).
  • AWS / EKS (advanced).
  • Linux (advanced).
  • Terraform and IaC in general (proficient).
  • Helm (proficient).
  • MongoDB (or similar).
  • Redis (or similar).
  • Monitoring – Prometheus, Grafana, Thanos (familiar).
  • Grasp of networking concepts (subnets, routing, peering, load balancing, NAT, etc.).
  • Common networking protocols (DNS, TCP/IP, HTTP, TLS, UDP).
  • Proactive, energetic, innovative and change-oriented.
  • A desire to lead/mentor a team.

Here’s why you should join us:

  • Everyone has unlimited paid holidays.
  • We have total flexibility in hours.
  • Employee share scheme.
  • Generous maternity and paternity leave.
  • Volunteering days.
  • Employee Wellbeing platform.

We all share the same vision – we value authenticity, respect, responsibility, independence, honesty, diversity and inclusion and most importantly treating others how you wish to be treated.

Our values include:

  • It’s ok to screw up!
  • The only stupid idea is the untested one!
  • Trust starts with you – make it count!
  • Assume best intent!
  • Make things better!

Tyk is an equal opportunities employer and we are determined to ensure that no applicant or employee receives less favourable treatment on the grounds of gender, age, disability, religion, belief, sexual orientation, marital status, or race.

Site Reliability Engineer - APAC in London employer: Tyk Technologies

Tyk is an exceptional employer that champions flexibility and autonomy, offering unlimited paid holidays and the ability to work remotely from anywhere in the world. With a strong commitment to employee growth, you will have the opportunity to develop your accounting skills while being part of a supportive team that values curiosity and innovation. Our inclusive work culture fosters collaboration and encourages you to take ownership of your role, making Tyk a truly rewarding place to build your career.

T

Contact Details:

Tyk Technologies Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Site Reliability Engineer - APAC in London

Tip Number 1

Network like a pro! Reach out to current or former employees at Tyk on LinkedIn. A friendly chat can give you insider info and maybe even a referral, which can really boost your chances.

Tip Number 2

Prepare for the interview by brushing up on your SRE skills. Be ready to discuss your experience with cloud technologies and how you've tackled reliability issues in the past. Show us you're the original thinker we’re looking for!

Tip Number 3

Don’t just wait for job openings; create your own opportunities! If you see a project or initiative at Tyk that excites you, mention it in your conversations. It shows initiative and genuine interest in our mission.

Tip Number 4

Apply through our website! It’s the best way to ensure your application gets seen by the right people. Plus, it shows you’re serious about joining our team and contributing to our vision.

We think you need these skills to ace Site Reliability Engineer - APAC in London

Site Reliability Engineering (SRE)
Cloud Technologies
SLA / SLO / SLI Management
Communication Skills
Leadership Skills
Operational Process Improvement
Software Design

Some tips for your application 🫡

Tailor Your Application:Make sure to customise your CV and cover letter for the Site Reliability Engineer role. Highlight your experience with cloud technologies, automation, and any relevant projects that showcase your skills. We want to see how you can contribute to our mission!

Show Off Your Technical Skills:Don’t hold back on showcasing your technical prowess! Mention your experience with Kubernetes, AWS, and any programming languages like Go or Python. We love seeing candidates who can demonstrate their hands-on experience in these areas.

Be Authentic:We value authenticity, so let your personality shine through in your application. Share your passion for SRE and how you approach problem-solving. We’re looking for original thinkers who aren’t afraid to challenge the status quo!

Apply Through Our Website:For the best chance of success, make sure to apply directly through our website. This way, we can easily track your application and get back to you quicker. Plus, it shows you’re serious about joining our team!

How to prepare for a job interview at Tyk Technologies

Know Your Stuff

Make sure you brush up on your knowledge of cloud technologies, especially AWS and Kubernetes. Be ready to discuss your experience with SLA/SLO/SLI management and how you've tackled reliability issues in the past.

Show Your Problem-Solving Skills

Prepare to share specific examples of how you've identified and resolved performance issues. Think about times when you conducted root-cause analysis and implemented solutions, as this will demonstrate your hands-on experience.

Be a Team Player

Tyk values collaboration, so be ready to talk about how you've worked with both technical and commercial teams. Highlight any experiences where you’ve led or mentored others, as well as how you’ve contributed to a positive team culture.

Ask Insightful Questions

Prepare thoughtful questions that show your interest in Tyk's mission and culture. Inquire about their approach to continuous improvement and how they handle incident response, as this will reflect your proactive mindset and alignment with their values.