At a Glance
- Tasks: Join us to enhance software operations and automate processes in a dynamic tech environment.
- Company: Reward Gateway, a leader in employee engagement and benefits, fostering innovation and collaboration.
- Benefits: Enjoy a competitive salary, hybrid work model, and opportunities for professional growth.
- Other info: Be part of a diverse team committed to making the world a better place to work.
- Why this job: Make a real impact by improving workplace experiences through cutting-edge technology.
- Qualifications: Experience in application support, PHP, Kubernetes, and strong problem-solving skills required.
The predicted salary is between 60000 - 80000 £ per year.
Reward Gateway, part of Edenred, is a global leader in benefits and employee engagement.
We help businesses attract, engage, and retain top talent through strategic reward, recognition, and well-being solutions.
Guided by our shared missions - 'Making the World a Better Place to Work' and 'Enriching Connections, For Good' - we’re committed to transforming workplaces and improving people’s daily lives.
Our team embodies entrepreneurial spirit, innovation, and respect.
We push boundaries, speak up, and stay human, fostering a culture where imagination thrives.
Role offers a hybrid work model to be present in our London office twice a week.
Your Role in our Mission
This hands-on role sits at the intersection of operational excellence and engineering craft.
You'll bridge the gap between traditional application support and software engineering by executing scripted remediation, configuration management, feature flag operations, safe, bounded code-level fixes, and runbook automation — all under clearly defined guardrails.
The goal is to reduce unnecessary L3 escalations while increasing autonomy, quality, and impact for our Application Operations function.
You'll apply these practices across our AWS environment (EKS), PHP services, and My SQL databases, using Datadog as our observability platform, Kibana for log exploration, and Heap to help quantify and understand customer impact.
Key Responsibilities
- Provide high-quality, timely L2.5 support for PHP applications running on EKS with My SQL backends, operating within clear guardrails that include configuration changes, feature flag operations, scripted runbooks, and safe, bounded code-level fixes.
- Model a shift-left mindset: resolve more at L2.5, automate more, and elevate less, increasing the percentage of incidents resolved without L3 involvement and improving MTTR.
- Participate in a healthy, sustainable on-call rotation with fair schedules, clear escalation paths, and strong post-incident learning practices.
- Apply engineering discipline to operational work: use version control, code review, and testing standards for scripts, runbooks, and automation tooling you produce.
- Develop and maintain automation scripts, runbooks, and playbooks for known issue patterns across workloads, services, and operational scenarios.
- Identify and automate repetitive remediation tasks to reduce manual toil and improve MTTR.
- Collaborate with peers to ensure the right monitoring signals, dashboards, and alerts exist in Datadog.
Tune app-level alerts and dashboards to minimize noise and surface actionable signals.
- Use Kibana to interrogate logs and correlate events with Datadog signals during investigations; improve log usefulness by feeding back patterns for better parsing and context.
- Use Heap to triangulate and quantify customer impact during incidents and problem investigations; incorporate findings into incident timelines and post-incident reviews.
- Participate in service onboarding and operability reviews to ensure new and changed services meet defined supportability standards before production.
- Contribute to the Service Catalogue with accurate ownership, SLAs/SLOs, runbooks, and escalation paths for supported services.
- Act as a first responder for application incidents at L2.5: triage, diagnose, and remediate within guardrails; support major incidents by providing technical context, structured diagnostics, Datadog/Kibana evidence, Heap impact analysis, and coordinated remediation alongside the incident commander.
- Use structured diagnostics before escalating — attach clear evidence, reproducibility steps, and impact assessments to every L3/SRE handoff.
- Feed operational findings into Problem Management and contribute to post-incident reviews; capture learning in improved runbooks, alerts, and automation.
- Help define, measure, and report on operational KPIs such as MTTR, percentage resolved at L2/L2.5, escalation rate, first-contact resolution, and SLO adherence.
- Continuously assess processes and workflows, delivering improvements that increase efficiency, consistency, and quality; balance reactive demand with proactive improvement work in Agile-aligned ways of working.
- Maintain high standards of documentation — runbooks, known errors, and operational guides are accurate, accessible, and kept up to date.
- Work closely with the Director of Application Operations, Problem Manager, and PETO peers (Platform, Infrastructure, Data, SRE) to ensure a coherent, joined-up operational approach.
- Partner with product-aligned engineering teams to understand application architecture, service dependencies, and failure modes; encode this knowledge into operational capabilities and runbooks.
- In scope: application-centric remediation under guardrails; automation of known issue patterns; high-quality runbooks; structured diagnostics; service readiness/documentation for PHP services on EKS with My SQL; ownership of app-level dashboards/alerts in Datadog, investigative use of Kibana logs, and customer-impact analysis via Heap.
- Working Hours and Practices
- Standard hours are 9am - 6pm, Mon - Fri.
- 1 day in every 4 is on call, paid at 1.5x hourly rate.
- On call hours are 6pm - 9am.
- If your on call day falls on a weekend, 24 hour on call cover is required.
Skills
- Proven experience in application support or operations engineering in cloud environments, ideally supporting PHP services running on Kubernetes (EKS) with My SQL backends.
- Hands‑on capability in at least one backend language (PHP preferred; Python or similar also valuable) sufficient to read, diagnose, and write safe operational scripts and minor fixes under guardrails.
- Practical Kubernetes skills for operations: kubectl/Helm basics, investigating pods/deployments, reading logs/events, understanding readiness/liveness probes, and performing safe rollouts/rollbacks within documented guardrails.
- My SQL operational fluency: connection and pool issues, slow query detection, query plan basics, common remediation patterns, and understanding of replication/backup implications.
- Strong experience using Datadog (APM/metrics/traces/dashboards/alerts) for investigation and detection; confident using Kibana for log exploration and correlation; ability to leverage Heap to assess user impact and prioritize remediation.
- Familiarity with ITSM tooling (e. g., Jira Service Management) and ITIL‑aligned incident and problem management processes.
- Strong communication skills; clear, concise documentation; collaborative approach focused on reducing toil, increasing automation, and raising the quality bar.
- Nice to Have Technical Skills
- Python.
- Experience with feature flag platforms and configuration-as-code within safe operational guardrails.
- Familiarity with AWS services that commonly interface with PHP/EKS workloads (e. g., Cloud Watch, ALB, S3, SQS) and how they surface in Datadog and Kibana.
- Exposure to service onboarding/operability reviews, SLOs, and contributing to a Service Catalogue.
- Experience balancing incident response with proactive improvement work in Agile contexts; strong documentation discipline.
- What Success Looks Like
- Increased percentage of tickets resolved at L2/L2.5, with reduced unnecessary L3 escalations.
- A maintained and actively used library of runbooks and automation scripts covering EKS/PHP/My SQL operational scenarios.
- Measurable reduction in MTTR driven by improved tooling, documentation, and automation.
- Earned trust of engineering and product peers as a technically credible, collaborative operations engineer.
- Company Commitment to Diversity
At Reward Gateway | Edenred we are committed to ensuring an inclusive and accessible recruitment process for all candidates.
If you have any specific requirements or need reasonable adjustments at any stage of the recruitment journey, please let your Talent Acquisition Partner know.
Your needs are important to us, and we want to ensure an equitable experience for every candidate.
About Our Values
Be comfortable.
Be you.
At Reward Gateway, we want all our employees to feel comfortable bringing their passion, creativity and individuality to work.
We value all cultures, backgrounds, and experiences, as we truly believe that diversity drives innovation.
Express yourself, join our community and help us Make the World a Better Place to Work.
- Third Floor, 1 Dean Street London W1D 3RB United Kingdom
- Engineering London Full Time £85,000 - £110,000 / year
- #J-18808-Ljbffr
Senior Software Engineer in London employer: Rewardgateway
Reward Gateway, part of Edenred, is an exceptional employer that prioritises employee well-being and engagement through a vibrant work culture and innovative benefits. With a flexible holiday plan, generous wellbeing allowance, and opportunities for professional development, employees are empowered to thrive both personally and professionally. The hybrid working model fosters a balanced lifestyle while collaborating with a diverse team dedicated to making workplaces better for everyone.
StudySmarter Expert Advice🤫
We think this is how you could land Senior Software Engineer in London
✨Join Local Tech Meetups
Get out there and mingle with fellow developers by joining local tech meetups. It’s a fantastic way to meet people who might be working at Rewardgateway or know someone who does. Plus, you can pick up some trendy tech skills and trends while you're at it!
✨Contribute to Open Source Projects
Show off your coding chops by jumping into open-source projects. Not only does this give you practical experience, but it also gets you noticed in the dev community. You'll create a killer portfolio that speaks volumes about your skills to Rewardgateway.
✨Tap into Online Developer Communities
Don’t underestimate the power of online developer communities like GitHub, Stack Overflow, and even Reddit. Participate in discussions, share your projects, and build your visibility. We can often find opportunities through these channels that can lead to a full-time gig at companies like Rewardgateway.
✨Explore Job Boards Specifically for Tech Roles
Keep your eyes peeled on job boards that focus on tech roles. Sites like TechCareers or Stack Overflow Jobs can often have listings for companies like Rewardgateway that might not show up on broader job sites. Make it a habit to check these regularly, and don’t hesitate to apply directly through our website!
We think you need these skills to ace Senior Software Engineer in London
Some tips for your application 🫡
Show off your coding skills:When applying for a software engineering role, it's super important to showcase your coding skills. Make sure your CV includes your tech stack, any relevant programming languages you’re comfortable with, and examples of projects you've worked on. If you have a GitHub profile, link it up! We love to see code in action.
Tailor your portfolio:For a full-time role, we’d expect to see some solid examples of your work in your portfolio. Make sure to include at least two or three projects that highlight your problem-solving skills and your ability to work with different technologies. Focus on the projects that are most relevant to the position at Rewardgateway.
Craft a killer cover letter:Your cover letter is your chance to stand out—make it personal! Explain why you want to work at Rewardgateway and how your skills align with the role. Show us your passion for software development. We dig enthusiastic candidates who understand the value of collaboration and continuous learning!
Be clear and concise:When it comes to writing your CV and cover letter, clarity is key. Avoid jargon that could confuse us and stick to simple, direct language. Highlight your achievements with quantifiable results where possible, and keep everything easy to read. A well-organised application goes a long way!
How to prepare for a job interview at Rewardgateway
✨Brush Up on Your Coding Skills
For a full-time software engineering role, it's crucial that we stay sharp with our coding abilities. Expect technical questions that might involve solving problems on the spot or discussing algorithms. Practise on platforms like LeetCode or HackerRank to get comfortable with the types of questions that often come up.
✨Know Your Tools and Frameworks
Make sure we’re well-acquainted with the tools and technologies listed in the job description. Familiarise ourselves with any specific frameworks or programming languages mentioned. If Rewardgateway uses React or Node.js, for instance, be ready to discuss how we’ve used them in previous projects or coursework.
✨Showcase Your Projects
Bring along a portfolio that highlights our best work. This could be code samples, GitHub repositories, or any side projects we’ve built. Make sure we can talk through our thought process for each project, especially the challenges we faced and how we solved them—this shows our problem-solving skills in action.
✨Prepare for Behavioural Questions
While technical skills are key, full-time positions also require cultural fit. Be ready to discuss our previous experiences and how we handle teamwork, conflict, and deadlines. Brush up on the STAR method—Situation, Task, Action, Result—to clearly articulate our past experiences when discussing how we've contributed to a team.