At a Glance
- Tasks: Lead reliability initiatives and enhance cloud infrastructure for GoDaddy's Commerce platform.
- Company: Join GoDaddy, a leader in empowering entrepreneurs with innovative online tools.
- Benefits: Enjoy remote work flexibility, competitive pay, health benefits, and generous time off.
- Other info: Diverse and inclusive culture that values your unique background and ideas.
- Why this job: Make a real impact on the future of reliability engineering while growing your career.
- Qualifications: 5+ years in site reliability, strong AWS and Linux skills, and a passion for automation.
The predicted salary is between 63000 - 77000 £ per year.
Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. Remote: This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings.
About The Team: The Commerce Site Reliability Engineering team is responsible for the reliability, scalability, and day-to-day operation of the platforms that power GoDaddy's Commerce ecosystem. We build and operate shared infrastructure, support critical production systems, and partner closely with engineering teams to ensure services remain secure, resilient, and highly available. As a Senior Site Reliability Engineer, you'll join a team that values ownership, operational excellence, and continuous improvement. Engineers are empowered to identify problems, drive meaningful change, and influence how reliability is delivered across the broader Commerce organisation. From improving operational maturity and reducing toil to modernising delivery platforms and strengthening incident response practices, this team plays a key role in enabling engineering teams to move quickly and safely. You'll work closely with engineers across infrastructure, cloud, security, networking, and application teams while helping shape the future of reliability engineering at GoDaddy. This role offers significant opportunity to broaden your impact, develop technical leadership skills, and grow toward Staff and Principal engineering positions over time.
What you'll get to do:
- Lead reliability and operational improvement initiatives across GoDaddy's Commerce platform, helping engineering teams build and operate services safely and at scale.
- Own critical production systems, drive incident response and post-incident improvements, and continuously raise the bar for operational excellence.
- Design, build, and enhance cloud infrastructure, automation, observability, and deployment platforms that improve reliability, scalability, and developer productivity.
- Partner with engineering, infrastructure, security, and product teams to solve complex technical challenges, manage operational risk, and support business-critical services.
- Use automation, AI-assisted engineering tools, and data-driven insights to reduce operational toil, improve diagnostics, and accelerate delivery.
- Mentor engineers, share knowledge, and influence engineering practices that improve reliability across the broader Commerce organisation.
- Contribute to the team's technical direction by identifying opportunities to improve systems, processes, and operational maturity.
Your experience should include:
- Significant experience (5 years +) operating, troubleshooting, and improving large-scale production systems in cloud-based environments.
- Strong expertise in AWS, Linux, container platforms such as Kubernetes, and modern infrastructure engineering practices.
- Experience building and maintaining Infrastructure as Code, automation solutions, and CI/CD pipelines that improve reliability, scalability, and delivery confidence.
- A proven track record of leading or owning production incidents, driving root cause analysis, and implementing long-term reliability improvements.
- Strong software engineering or scripting skills using languages such as Python, Go, TypeScript, or similar technologies.
- Experience using observability data, monitoring, and operational metrics to identify issues, improve system performance, and support data-driven decision making.
- Demonstrated ability to independently lead complex technical initiatives, manage competing priorities, and deliver outcomes across multiple teams or stakeholders.
- Experience mentoring engineers, influencing technical decisions, and helping raise operational and engineering standards within a team.
- A continuous improvement mindset with a focus on automation, reducing operational toil, and leaving systems and processes better than you found them.
- Experience using AI-assisted engineering tools to improve productivity, accelerate troubleshooting, automate repetitive tasks, and enhance operational workflows.
You might also have:
- Experience with SaltStack, Ansible, Terraform, Pulumi, CloudFormation, or AWS CDK.
- Experience operating shared or multi-tenant infrastructure services at scale.
- Experience supporting eCommerce, payments, fintech, or other high-availability customer-facing platforms.
- Experience designing observability, SLO, SLI, or operational readiness frameworks.
- Experience contributing to architectural strategy and roadmap planning.
- Experience leveraging AI-assisted tools to improve engineering productivity and operational effectiveness.
We encourage you to apply even if your experience or skillset doesn’t align perfectly with every requirement. We value a wide range of backgrounds and transferable skills, and we are excited to support learning and growth.
We've got your back: We offer a range of total rewards that may include paid time off, retirement savings (e.g., 401k, pension schemes), bonus/incentive eligibility, equity grants, participation in our employee stock purchase plan, competitive health benefits, and other family-friendly benefits including parental leave. GoDaddy’s benefits vary based on individual role and location and can be reviewed in more detail during the interview process.
About us: GoDaddy is empowering everyday entrepreneurs around the world by providing the help and tools to succeed online, making opportunity more inclusive for all. GoDaddy is the place people come to name their idea, build a professional website, attract customers, sell their products and services, and manage their work. Our mission is to give our customers the tools, insights, and people to transform their ideas and personal initiative into success. To learn more about the company, visit About Us. At GoDaddy, we know diverse teams build better products—period. Our people and culture reflect and celebrate that sense of diversity and inclusion in ideas, experiences, and perspectives. But we also know that’s not enough to build true equity and belonging in our communities. That’s why we prioritise integrating diversity, equity, inclusion, and belonging principles into the core of how we work every day—focusing not only on our employee experience but also our customer experience and operations. It’s the best way to serve our mission of empowering entrepreneurs everywhere and making opportunity more inclusive for all.
GoDaddy is proud to be an equal opportunity employer. GoDaddy will consider for employment qualified applicants with criminal histories in a manner consistent with local and federal requirements.
Senior Site Reliability Engineer employer: Go Daddy Group
At GoDaddy, we pride ourselves on being an exceptional employer that champions a culture of innovation and inclusivity. As a Senior Site Reliability Engineer, you'll enjoy the flexibility of remote work while collaborating with diverse teams dedicated to operational excellence and continuous improvement. With ample opportunities for professional growth, competitive benefits, and a commitment to empowering entrepreneurs, GoDaddy is the ideal place for those seeking meaningful and rewarding careers.