Director of Site Reliability Engineering

Director of Site Reliability Engineering

Full-Time 63000 - 77000 Β£ / year (est.) No working from home possible
Hackajob Ltd

At a Glance

  • Tasks: Lead innovative projects to enhance software reliability and operational efficiency.
  • Company: Join JPMorgan Chase, a global leader in financial services.
  • Benefits: Competitive salary, comprehensive health coverage, and generous perks like tuition reimbursement.
  • Other info: Diverse and inclusive workplace with excellent career growth opportunities.
  • Why this job: Make a significant impact in a dynamic environment with cutting-edge technology.
  • Qualifications: 10+ years in site reliability engineering with advanced technical skills.

The predicted salary is between 63000 - 77000 Β£ per year.

Join a globally recognized financial organization and advance your profession to new heights by contributing to revolutionary projects. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology and Enterprise Technology Team, you will draw upon your advanced knowledge to identify new opportunities to influence critical incident management and improve the end-to-end lifecycle of software development for the firm. You will have the opportunity to manage, design, and implement infrastructure components to improve reliability and ensure operational efficiency.

Job Responsibilities:

  • Identifies and solves problems of high complexity and drives improvements as outcomes.
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate complex incident analysis and reliability decisioning, validating outputs and handling operational data according to sensitivity and security requirements.
  • Works with development teams throughout the software life cycle ensuring sustainable software releases, leading medium to large projects by bringing together the proper perspectives, identifying road blockers, and integrating feedback from team members and the firm.
  • Manages, designs, and implements infrastructure components to improve reliability and ensure operational efficiency.
  • Influences management and stakeholders through clear recommendations and measurable reliability outcomes.
  • Leads reuse-first adoption of AI-assisted reliability workflows across SDLC/toolchain practices, ensuring traceability/auditability, resiliency, and security controls.

Required qualifications, capabilities, and skills:

  • Formal training or certification on site reliability engineering concepts and 10+ years applied experience.
  • Advanced knowledge of software applications and technical processes with considerable depth in one or more technical disciplines.
  • Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve reliability engineering workflows with strong validation habits and awareness of data sensitivity.
  • Ability to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements.
  • Ability to determine how each system relates to each other and uses a breadth of tools to drive automation, process improvement, and reliability for the firm.
  • Experience with translating research, analysis, and tests into business recommendations.

Preferred qualifications, capabilities, and skills:

  • Experience modernizing reliability practices in large, multi-team environments and/or legacy-to-modern transformations.
  • Experience building "paved road" reliability capabilities (shared libraries, templates, tooling) that scale across many teams.
  • Familiarity with SLO programs and operational readiness practices at scale.

We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more.

We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law.

Director of Site Reliability Engineering employer: Hackajob Ltd

At loveholidays, we pride ourselves on fostering a collaborative and innovative work culture that empowers our employees to thrive. As a Product Designer, you'll have the opportunity to contribute to meaningful projects that enhance customer experiences while enjoying a range of benefits, including professional development opportunities and a supportive team environment in a vibrant location. Join us in our mission to make travel accessible for everyone and be part of a company that values your creativity and input.

Hackajob Ltd

Contact Details:

Hackajob Ltd Recruitment Team

We think you need these skills to ace Director of Site Reliability Engineering

Site Reliability Engineering
Incident Management
Infrastructure Design and Implementation
AI Capabilities in Reliability Engineering
Software Development Life Cycle (SDLC)
Automation and Process Improvement
Data Sensitivity Awareness