Vice President, Devops Production Services

Vice President, Devops Production Services

Full-Time 63000 - 77000 Β£ / year (est.) No working from home possible
Hackajob Ltd

At a Glance

  • Tasks: Manage and support critical AI applications in a fast-paced production environment.
  • Company: Join BNY Mellon, a leading global financial services company.
  • Benefits: Competitive salary, growth opportunities, and a culture of innovation.
  • Other info: Dynamic team environment with opportunities for continuous improvement and learning.
  • Why this job: Make a real impact with cutting-edge technology in finance.
  • Qualifications: Experience in production application support and strong SQL skills required.

The predicted salary is between 63000 - 77000 Β£ per year.

hackajob is partnering directly with BNY Mellon to hire for this role. This role is located in Manchester.

Role Overview: We are seeking a highly skilled professional with strong experience in Production Application Support to manage and support critical enterprise AI based applications in a fast-paced production environment. The role requires hands-on expertise in monitoring, incident management, troubleshooting, release support, and ensuring high availability and stability of business-critical platforms.

In this role, you'll make an impact in the following ways:

  • Provide L2/L3 production support for enterprise applications and ensure platform stability, resiliency, and availability.
  • Monitor application health, system performance, batch jobs, interfaces, and alerts using enterprise monitoring and observability tools.
  • Investigate, troubleshoot, and resolve production incidents within defined SLAs.
  • Perform root cause analysis (RCA) for recurring issues and drive permanent fixes.
  • Analyze production logs, identify failure patterns, and create actionable dashboards to improve service monitoring and incident response.
  • Coordinate with development, infrastructure, database, network, and business teams for issue resolution.
  • Support application deployments, change requests, weekend releases, and post-release validations.
  • Maintain incident, problem, and change records in service management tools.
  • Drive continuous service improvement through automation, process optimization, and proactive monitoring.
  • Participate in on-call support and major incident management as required.
  • Prepare operational reports, service health summaries, and stakeholder communications.
  • Write and analyze SQL queries for data validation, issue investigation, and production troubleshooting.
  • Use Unix/Linux commands and scripting for application support, log reviews, file handling, and system-level troubleshooting.
  • Leverage Splunk extensively for log analysis, issue diagnosis, trend identification, alerting insights, and dashboard creation.

To be successful in this role, we're seeking the following:

  • Proven experience in production application support for business-critical applications.
  • Strong understanding of incident management, problem management, and change management processes.
  • Strong SQL skills for querying, troubleshooting, and data analysis in production environments.
  • Extensive hands-on experience with Splunk for log analysis, search creation, troubleshooting, monitoring, and dashboard development.
  • Strong Unix/Linux skills for navigating servers, reviewing logs, troubleshooting jobs/processes, and supporting application runtime environments.
  • Experience with monitoring and alerting tools, log analysis, Grafana, and dashboard-based production support.
  • Experience with ITSM tools such as ServiceNow, Jira, or similar platforms.
  • Ability to analyze application, infrastructure, and integration issues across distributed systems.
  • Experience supporting applications in cloud and/or on-prem environments.
  • Familiarity with scripting and troubleshooting middleware/interfaces.
  • Strong knowledge of release support, service recovery, and operational governance.
  • Ability to work in a high-pressure environment with strong ownership and accountability.
  • Demonstrated ability to ramp up quickly on new applications, platforms, and support processes, with strong learning agility and immediate contribution in a fast-paced production environment.

Much Preferred Skills:

  • Azure Cloud experience preferred.
  • Knowledge of automation/scripting using Python, Shell, or PowerShell.
  • Exposure to DevOps / SRE practices, CI/CD pipelines, and observability tooling.
  • Strong communication skills with the ability to provide concise incident and executive status updates.

At BNY, our culture allows us to run our company better and enables employees' growth and success. As a leading global financial services company at the heart of the global financial system, we influence nearly 20% of the world's investible assets. Every day, our teams harness cutting-edge AI and breakthrough technologies to collaborate with clients, driving transformative solutions that redefine industries and uplift communities worldwide. Recognized as a top destination for innovators, BNY is where bold ideas meet advanced technology and exceptional talent. Together, we power the future of finance - and this is what #LifeAtBNY is all about. Join us and be part of something extraordinary.

Vice President, Devops Production Services employer: Hackajob Ltd

At loveholidays, we pride ourselves on fostering a collaborative and innovative work culture that empowers our employees to thrive. As a Product Designer, you'll have the opportunity to contribute to meaningful projects that enhance customer experiences while enjoying a range of benefits, including professional development opportunities and a supportive team environment in a vibrant location. Join us in our mission to make travel accessible for everyone and be part of a company that values your creativity and input.

Hackajob Ltd

Contact Details:

Hackajob Ltd Recruitment Team

We think you need these skills to ace Vice President, Devops Production Services

Production Application Support
Incident Management
Troubleshooting
Release Support
Monitoring and Observability Tools
Root Cause Analysis (RCA)
SQL Skills