Lead, Site Reliability Engineer (Infrastructure operations) in Dundrum

Lead, Site Reliability Engineer (Infrastructure operations) in Dundrum

Dundrum Full-Time No working from home possible
M

At a Glance

  • Tasks: Lead the reliability of Mastercard's critical payment systems and enhance service quality.
  • Company: Join Mastercard, a global leader in digital payments and innovation.
  • Benefits: Competitive salary, inclusive culture, and opportunities for professional growth.
  • Other info: Dynamic team environment with on-call responsibilities and continuous learning opportunities.
  • Why this job: Make a real impact on global transactions while working with cutting-edge technology.
  • Qualifications: 5-10 years in SRE or related roles, strong problem-solving skills, and experience with monitoring tools.

Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.

About the Role: Mastercard’s Program aligned Site Reliability Engineering (SRE) teams are dedicated to delivering a seamless experience for our customers. We achieve this by maintaining every aspect of our Programs infrastructure and technology ecosystem to the highest standards, ensuring compliance with rigorous security requirements. Within Mastercard, SRE focuses on the reliability and performance of core infrastructure, networks, and foundational services that power our applications. Our mission is to ensure these components operate with excellence, enabling applications to deliver an outstanding customer experience. In this role, you will join our Payments Network SRE team and take ownership of continuously assessing and elevating the end to end service quality of our platform. You will leverage data to drive root cause analysis and deliver strategic insights to key stakeholders on resource utilization, capacity forecasting, and performance trends—ensuring the availability, scalability, and resilience of our network.

Key Responsibilities:

  • Lead continuous assessments of the application infrastructure supporting critical Mastercard applications, focusing on health, performance, monitoring and alerting, and capacity analysis.
  • Collaborate with Product and Development teams to forecast growth requirements and ensure scalability and resiliency.
  • Champion observability as a core principle for infrastructure services by assessing environments and technologies to uncover gaps in monitoring and alerting.
  • Design and implement strategies to close these gaps, ensuring all infrastructure telemetry is integrated into a unified, single-pane-of-glass view.
  • Build custom dashboards to investigate and perform root cause analysis on complex issues.
  • Lead regular incident reviews with internal support teams to ensure root causes are identified.
  • When patterns of failure or compatibility issues between software and infrastructure emerge, develop and implement strategies to remediate or mitigate risks.
  • Leverage automation and AI technologies to enhance proactive issue detection, enable self-healing capabilities, reducing Mean Time to Detect (MTTD) and Mean Time to Mitigate (MTTM).
  • Develop testing and validation plans for new environment builds, disaster recovery exercises and post-maintenance activities to certify environment readiness before customer traffic is routed to it.
  • Champion continuous learning, development, and knowledge sharing across networking and other infrastructure disciplines to strengthen multi-disciplinary SRE team capabilities.
  • Lead training initiatives for team members and Product and Development on networking aspects of the platforms.
  • Evaluate vendor hardware, firmware, and software upgrade roadmaps, and conduct proof-of-concept (POC) testing to identify potential risks and opportunities for improvement in upcoming releases.

All about you:

  • 5–10 years of experience in an SRE or SRE related operations role, including 3+ years supporting e-commerce, financial services, or large scale SaaS platforms.
  • Excellent infrastructure troubleshooting and analytical problem solving skills.
  • Strong hands on experience with observability and monitoring tools such as Splunk, Dynatrace, or equivalent, with a proven ability to triage and investigate complex issues.
  • Familiarity with network telemetry tools such as SolarWinds and NetScout.
  • Proficiency in packet level debugging, including capturing traffic with tools like tcpdump and analyzing packets using Wireshark.
  • Broad understanding of end to end infrastructure supporting payment platforms—spanning platform services, networking, databases, and storage.
  • Experience with automation and Infrastructure as Code tools such as Chef, Ansible, and Terraform, as well as structured data formats (JSON/YAML).
  • Excellent communication skills with the ability to coordinate cross functional troubleshooting efforts and lead RCA processes to closure.
  • Demonstrated ability to troubleshoot complex production issues, perform root cause analysis, and drive long term corrective actions.
  • Experience partnering with development teams to shape architecture, define SLIs/SLOs, and embed reliability into services from design through operation.
  • Strong understanding of monitoring and observability ecosystems, including Prometheus, Grafana, ELK/EFK, Splunk, Dynatrace, and OpenTelemetry.
  • Effective incident management skills with a structured, analytical approach to problem solving.

The Payments Network SRE team is responsible for the runtime availability of some of Mastercard’s most critical core payment systems, which support national infrastructure and operate 24/7 year-round. As a result, this role will include periodic on-call responsibilities when required.

Corporate Security Responsibility

All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:

  • Abide by Mastercard’s security policies and practices;
  • Ensure the confidentiality and integrity of the information being accessed;
  • Report any suspected information security violation or breach, and
  • Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines.

Lead, Site Reliability Engineer (Infrastructure operations) in Dundrum employer: Mastercard

As a Senior Site Reliability Engineer at Connex, you will be part of a dynamic team dedicated to maintaining the UK's national payment infrastructure, where innovation and collaboration are at the forefront of our work culture. We offer a supportive environment that prioritises continuous improvement and professional growth, alongside competitive benefits and a commitment to work-life balance. Join us in making a meaningful impact in the financial services sector while enjoying the unique advantages of working in a cutting-edge technology space.

M

Contact Details:

Mastercard Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Lead, Site Reliability Engineer (Infrastructure operations) in Dundrum

Join Local Tech Meetups

Get out there and mingle with fellow developers by joining local tech meetups. It’s a fantastic way to meet people who might be working at Mastercard or know someone who does. Plus, you can pick up some trendy tech skills and trends while you're at it!

Contribute to Open Source Projects

Show off your coding chops by jumping into open-source projects. Not only does this give you practical experience, but it also gets you noticed in the dev community. You'll create a killer portfolio that speaks volumes about your skills to Mastercard.

Tap into Online Developer Communities

Don’t underestimate the power of online developer communities like GitHub, Stack Overflow, and even Reddit. Participate in discussions, share your projects, and build your visibility. We can often find opportunities through these channels that can lead to a full-time gig at companies like Mastercard.

Explore Job Boards Specifically for Tech Roles

Keep your eyes peeled on job boards that focus on tech roles. Sites like TechCareers or Stack Overflow Jobs can often have listings for companies like Mastercard that might not show up on broader job sites. Make it a habit to check these regularly, and don’t hesitate to apply directly through our website!

We think you need these skills to ace Lead, Site Reliability Engineer (Infrastructure operations) in Dundrum

Site Reliability Engineering (SRE)
Infrastructure Operations
Observability and Monitoring Tools
Splunk
Dynatrace
Network Telemetry Tools
SolarWinds

Some tips for your application 🫡

Show off your coding skills:When applying for a software engineering role, it's super important to showcase your coding skills. Make sure your CV includes your tech stack, any relevant programming languages you’re comfortable with, and examples of projects you've worked on. If you have a GitHub profile, link it up! We love to see code in action.

Tailor your portfolio:For a full-time role, we’d expect to see some solid examples of your work in your portfolio. Make sure to include at least two or three projects that highlight your problem-solving skills and your ability to work with different technologies. Focus on the projects that are most relevant to the position at Mastercard.

Craft a killer cover letter:Your cover letter is your chance to stand out—make it personal! Explain why you want to work at Mastercard and how your skills align with the role. Show us your passion for software development. We dig enthusiastic candidates who understand the value of collaboration and continuous learning!

Be clear and concise:When it comes to writing your CV and cover letter, clarity is key. Avoid jargon that could confuse us and stick to simple, direct language. Highlight your achievements with quantifiable results where possible, and keep everything easy to read. A well-organised application goes a long way!

How to prepare for a job interview at Mastercard

Brush Up on Your Coding Skills

For a full-time software engineering role, it's crucial that we stay sharp with our coding abilities. Expect technical questions that might involve solving problems on the spot or discussing algorithms. Practise on platforms like LeetCode or HackerRank to get comfortable with the types of questions that often come up.

Know Your Tools and Frameworks

Make sure we’re well-acquainted with the tools and technologies listed in the job description. Familiarise ourselves with any specific frameworks or programming languages mentioned. If Mastercard uses React or Node.js, for instance, be ready to discuss how we’ve used them in previous projects or coursework.

Showcase Your Projects

Bring along a portfolio that highlights our best work. This could be code samples, GitHub repositories, or any side projects we’ve built. Make sure we can talk through our thought process for each project, especially the challenges we faced and how we solved them—this shows our problem-solving skills in action.

Prepare for Behavioural Questions

While technical skills are key, full-time positions also require cultural fit. Be ready to discuss our previous experiences and how we handle teamwork, conflict, and deadlines. Brush up on the STAR method—Situation, Task, Action, Result—to clearly articulate our past experiences when discussing how we've contributed to a team.