At a Glance
- Tasks: Solve complex production challenges and improve live operational environments with automation.
- Company: Dynamic tech company in London offering a hybrid work model.
- Benefits: Competitive salary, on-call allowance, and opportunities for professional growth.
- Other info: Enjoy a supportive environment with low call-out volumes and a focus on continuous improvement.
- Why this job: Join a team that values innovation and make a real impact on operational excellence.
- Qualifications: Experience in application operations and strong troubleshooting skills required.
The predicted salary is between 63000 - 77000 £ per year.
- Senior Software Engineer – Permanent – London/Hybrid - £70,000 – 85,000
- 2 days per week on site in central London, 3 days remote
This is a hands-on engineering role for people who enjoy solving complex production challenges and applying strong software engineering principles to improve live operational environments.
You will join a team of highly capable engineers who combine technical expertise, curiosity and analytical thinking to strengthen reliability, increase automation and prevent recurring issues.
By reducing repetitive operational work, the team is able to focus on higher-value engineering challenges across resilience, operational excellence and continuous improvement.
You will play a key role in eliminating operational toil, automating manual processes and improving service reliability, working closely with Software Engineering, Platform and Site Reliability Engineering teams across the organisation.
Core technologies
AWS (EKS/Kubernetes), PHP, My SQL, Datadog, Kibana and Heap.
This role offers the opportunity to
- Work on challenging technical problems across a modern cloud platform.
- Build automation and tooling that delivers measurable business impact.
- Develop expertise across software engineering, cloud operations and reliability engineering.
- Join a team that values technical excellence, ownership, collaboration and continuous improvement.
What You'll Do
- Resolve complex application and platform issues within clearly defined engineering guardrails.
- Diagnose production incidents, perform safe remediation and reduce unnecessary escalation to L3 teams.
- Build automation, scripts, runbooks and tooling that improve operational efficiency and reduce MTTR.
- Use Datadog, Kibana and Heap to investigate issues, understand customer impact and improve service health.
- Partner with Software Engineering, SRE and Platform teams to improve observability, reliability and operational readiness.
- Contribute to service onboarding, post-incident reviews and continuous improvement initiatives.
What We're Looking For
- Experience in Application Operations, Production Engineering, SRE or a similar operational engineering environment.
- Experience supporting cloud-hosted applications, ideally PHP services running on Kubernetes (EKS) with My SQL databases.
- Strong troubleshooting and diagnostic skills, with the ability to understand systems end-to-end.
- Hands‑on experience with PHP, Python or another backend language.
- Practical Kubernetes knowledge, including deployments, logging, monitoring and safe rollbacks.
- My SQL operational experience, including performance investigation and query analysis.
- Experience with observability platforms such as Datadog and log analysis tools such as Kibana.
- Strong communication skills and a passion for automation, continuous improvement and engineering quality.
- Nice to Have
- Feature flag platforms and configuration-as-code.
- AWS services such as S3, SQS, ALB and Cloud Watch.
- Experience with SLOs, service onboarding and operational readiness.
- Working Pattern & On-Call
- Standard hours: Monday to Friday, 9am to 6pm.
- 1-in-4 on-call rotation.
- £500 on-call allowance per week.
- Additional hourly payments for call-outs.
- Enhanced rates for public holiday coverage.
- Historically low call-out volumes, supported by a strong focus on automation, operational maturity and continuous improvement.
- #J-18808-Ljbffr
Senior Software Engineer - Operational Support employer: Robson Bale
Robson Bale is an excellent employer that prioritises employee growth and development, offering ongoing training and leadership opportunities for its team members. With a strong focus on collaboration and customer satisfaction, the work culture fosters innovation and support, making it an ideal environment for experienced IT professionals looking to make a meaningful impact in the telecommunications sector.