Systems Engineering Manager, Site Reliability Engineering, ML Compute in London

Systems Engineering Manager, Site Reliability Engineering, ML Compute in London

London Full-Time 63000 - 77000 £ / year (est.) No working from home possible
G

At a Glance

  • Tasks: Lead a team to ensure uptime and performance of key services while automating solutions.
  • Company: Join Google’s innovative Site Reliability Engineering team with a culture of curiosity and collaboration.
  • Benefits: Competitive salary, mentorship opportunities, and a dynamic work environment.
  • Other info: Opportunity for career growth in a supportive, blame-free environment.
  • Why this job: Make a real impact on large-scale systems and enhance user experiences globally.
  • Qualifications: Bachelor's degree in Computer Science, 5 years programming experience, and 3 years in management.

The predicted salary is between 63000 - 77000 £ per year.

info_outline

XIn most instances, this position requires in-person interviews as part of the hiring process.

Minimum qualifications

Bachelor's degree in Computer Science or a related technical field or equivalent practical experience.

  • 5 years of experience with programming in one or more programming languages.
  • 3 years of people management experience.
  • 3 years of experience leading projects and working with administration (e. g., filesystems, inodes, system calls) or networking (e. g., TCP/IP, routing, network topologies and hardware, SDN).

Preferred qualifications

  • Master's degree in Computer Science or related technical field involving coding (e. g., physics or mathematics).
  • Track record of mentoring technical leads.
  • Proven success leading and influencing multiple technical teams.

About the job

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems.

SRE ensures that Google's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement.

Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance.

Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating work through automation.

On the SRE team, you’ll have the opportunity to manage the complex challenges of scale which are unique to Google, while using your expertise in coding, algorithms, complexity analysis and large-scale system design.

SRE's culture of intellectual curiosity, problem solving and openness is key to its success.

Our organization brings together people with a wide variety of backgrounds, experiences and perspectives.

We encourage them to collaborate, think big and take risks in a blame-free environment.

We promote self-direction to work on meaningful projects, while we also strive to create an environment that provides the support and mentorship needed to learn and grow.

To learn more

check out our books on

Site Reliability Engineering or read a career profile about why a Software Engineer chose to join SRE.

The ML Compute SRE team mission is to deliver a ML compute infrastructure for all users.

We ensure that all ML accelerators (TPUs and GPUs) are fully and appropriately supported as part of the TI and Cloud Compute platforms and that all ML jobs run efficiently, safely and reliably.

We support both the hardware and the low level services that provide ML as an Iaa S.

Behind everything our users see online is the architecture built by the Technical Infrastructure team to keep it running.

From developing and maintaining our data centers to building the next generation of Google platforms, we make Google's product portfolio possible.

We're proud to be our engineers' engineers and love voiding warranties by taking things apart so we can rebuild them.

We keep our networks up and running, ensuring our users have the best and fastest experience possible.

Responsibilities

  • Lead a team of software/systems engineers on projects for users and be directly responsible for uptime.
  • Own end-to-end availability and performance of key services and build automation to prevent problem recurrence. Automate response to all non-exceptional service conditions.
  • Lead by example, mentor the team and establish credibility through quality technical execution.
  • Manage on-call rotations across continents, using a follow-the-sun model.
  • Design, write and deliver software to improve the availability, scalability, latency and efficiency of Google's services.

Systems Engineering Manager, Site Reliability Engineering, ML Compute in London employer: Google

As a Senior Manager in Ads Solutions Engineering at gTech, you will thrive in a dynamic and innovative environment that prioritises collaboration and professional growth. The company fosters a culture of continuous learning and development, offering ample opportunities to lead transformative projects while working with cutting-edge technologies. Located in a vibrant tech hub, gTech provides a unique chance to engage with top-tier clients and contribute to impactful solutions that drive success.

G

Contact Details:

Google Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Systems Engineering Manager, Site Reliability Engineering, ML Compute in London

Join Local Tech Meetups

Get out there and mingle with fellow developers by joining local tech meetups. It’s a fantastic way to meet people who might be working at Google or know someone who does. Plus, you can pick up some trendy tech skills and trends while you're at it!

Contribute to Open Source Projects

Show off your coding chops by jumping into open-source projects. Not only does this give you practical experience, but it also gets you noticed in the dev community. You'll create a killer portfolio that speaks volumes about your skills to Google.

Tap into Online Developer Communities

Don’t underestimate the power of online developer communities like GitHub, Stack Overflow, and even Reddit. Participate in discussions, share your projects, and build your visibility. We can often find opportunities through these channels that can lead to a full-time gig at companies like Google.

Explore Job Boards Specifically for Tech Roles

Keep your eyes peeled on job boards that focus on tech roles. Sites like TechCareers or Stack Overflow Jobs can often have listings for companies like Google that might not show up on broader job sites. Make it a habit to check these regularly, and don’t hesitate to apply directly through our website!

We think you need these skills to ace Systems Engineering Manager, Site Reliability Engineering, ML Compute in London

Programming Languages
People Management
Project Leadership
Filesystems
System Calls
Networking
TCP/IP

Some tips for your application 🫡

Show off your coding skills:When applying for a software engineering role, it's super important to showcase your coding skills. Make sure your CV includes your tech stack, any relevant programming languages you’re comfortable with, and examples of projects you've worked on. If you have a GitHub profile, link it up! We love to see code in action.

Tailor your portfolio:For a full-time role, we’d expect to see some solid examples of your work in your portfolio. Make sure to include at least two or three projects that highlight your problem-solving skills and your ability to work with different technologies. Focus on the projects that are most relevant to the position at Google.

Craft a killer cover letter:Your cover letter is your chance to stand out—make it personal! Explain why you want to work at Google and how your skills align with the role. Show us your passion for software development. We dig enthusiastic candidates who understand the value of collaboration and continuous learning!

Be clear and concise:When it comes to writing your CV and cover letter, clarity is key. Avoid jargon that could confuse us and stick to simple, direct language. Highlight your achievements with quantifiable results where possible, and keep everything easy to read. A well-organised application goes a long way!

How to prepare for a job interview at Google

Brush Up on Your Coding Skills

For a full-time software engineering role, it's crucial that we stay sharp with our coding abilities. Expect technical questions that might involve solving problems on the spot or discussing algorithms. Practise on platforms like LeetCode or HackerRank to get comfortable with the types of questions that often come up.

Know Your Tools and Frameworks

Make sure we’re well-acquainted with the tools and technologies listed in the job description. Familiarise ourselves with any specific frameworks or programming languages mentioned. If Google uses React or Node.js, for instance, be ready to discuss how we’ve used them in previous projects or coursework.

Showcase Your Projects

Bring along a portfolio that highlights our best work. This could be code samples, GitHub repositories, or any side projects we’ve built. Make sure we can talk through our thought process for each project, especially the challenges we faced and how we solved them—this shows our problem-solving skills in action.

Prepare for Behavioural Questions

While technical skills are key, full-time positions also require cultural fit. Be ready to discuss our previous experiences and how we handle teamwork, conflict, and deadlines. Brush up on the STAR method—Situation, Task, Action, Result—to clearly articulate our past experiences when discussing how we've contributed to a team.