At a Glance
- Tasks: Lead a team to design and maintain reliable digital services that drive economic growth.
- Company: Join Inspire People, a forward-thinking company focused on secure online services.
- Benefits: Competitive salary, professional development, and a collaborative work environment.
- Other info: Flexible work environment with opportunities for travel and team collaboration.
- Why this job: Make a real impact by enhancing system reliability and mentoring future engineers.
- Qualifications: 5+ years in Site Reliability Engineering with strong leadership and cloud expertise.
The predicted salary is between 63000 - 77000 Β£ per year.
Lead Site Reliability Engineer at Inspire People.
About the role Become a vital member of a committed digital team that is dedicated to delivering reliable, secure, and scalable online services aimed at promoting economic growth across the UK.
In partnership with Inspire People, the Department for Business and Trade (DBT) is on the lookout for a Lead Site Reliability Engineer who possesses a strong background in nurturing and guiding engineering teams.
This position calls for a solid foundation in Dev Ops and Site Reliability Engineering, along with expertise in cloud technologies, infrastructure-as-code practices, and a dedication to building resilient distributed systems.
- Key facts
- Location: South West London, London
- Engagement: Permanent
- Compensation: $80,000
- Team: Not specified
What you'll do
- Lead a talented group of engineers in the design, development, and maintenance of dependable digital services that meet user needs.
- Establish and implement industry best practices in Dev Ops and Site Reliability Engineering to optimize system performance and enhance reliability.
- Manage cloud infrastructure effectively, ensuring the application of infrastructure-as-code methodologies for streamlined operations.
- Cultivate a culture of continuous improvement and innovation within the engineering team, encouraging team members to share ideas and solutions.
- Collaborate with cross-functional teams to ensure alignment on project goals and deliverables, facilitating smooth communication and cooperation.
- Monitor system performance and reliability, utilizing appropriate tools to identify and resolve issues proactively.
- Develop and maintain documentation related to system architecture, processes, and procedures to ensure clarity and consistency.
- Mentor and guide junior engineers, fostering their professional growth and enhancing team capabilities.
Requirements
- At least 5 years of experience in Site Reliability Engineering or a closely related field, demonstrating a strong technical foundation.
- Proven experience in leadership roles, with a track record of managing and developing engineering teams effectively.
- In-depth knowledge of cloud platforms, particularly Amazon Web Services (AWS) or Microsoft Azure, and their associated services.
- Expertise in infrastructure-as-code tools such as Terraform or AWS Cloud Formation, enabling efficient infrastructure management.
- Familiarity with monitoring and incident response tools, ensuring quick identification and resolution of system issues.
- Comprehensive understanding of distributed systems and microservices architecture, with the ability to design and implement scalable solutions.
- Nice to have
- Experience with container orchestration technologies like Kubernetes, enhancing deployment and management of applications.
- Knowledge of implementing Continuous Integration and Continuous Deployment (CI/CD) pipelines to streamline development workflows.
- Awareness of security best practices in cloud environments, contributing to the overall safety and integrity of systems.
Skills & tools
- Proficient in AWS or Azure cloud services, leveraging their capabilities for effective service delivery.
- Skilled in using Terraform or Cloud Formation for infrastructure management and automation.
- Familiar with monitoring tools such as Prometheus and Grafana for performance tracking and alerting.
- Experienced in incident management systems to handle and resolve operational issues efficiently.
- Competent in scripting languages like Python or Bash to automate tasks and improve processes.
- Practical notes
- Please note that visa sponsorship is not available for this role, and applicants must have the right to work in the UK.
- Occasional travel may be required for team meetings or project-related activities, so flexibility is important.
- The position offers a competitive salary, along with opportunities for professional development and a supportive work environment that values collaboration and innovation.
- #J-18808-Ljbffr
Lead Site Reliability Engineer in Hounslow employer: Inspire People
At DVSA, we pride ourselves on being an excellent employer, offering a dynamic work environment where your expertise as a Senior Data Architect will directly influence data management and governance across our agency. With competitive salaries, generous leave policies, and robust career development opportunities, we foster a culture of collaboration and innovation, ensuring that our employees feel valued and empowered to make a meaningful impact on road safety in Great Britain.