Site Reliability Engineer – NS London

Site Reliability Engineer – NS London

London Full-Time 50000 - 70000 £ / year (est.) No working from home possible

At a Glance

  • Tasks: Automate system support and enhance reliability for national security applications.
  • Company: BAE Systems Digital Intelligence, a leader in tech innovation.
  • Benefits: Competitive salary, overtime benefits, and on-call allowances.
  • Other info: Be part of a collaborative DevOps/SRE community with growth opportunities.
  • Why this job: Join a dynamic team improving system health and making a real impact.
  • Qualifications: Experience in web development, Linux, and monitoring large systems.

The predicted salary is between 50000 - 70000 £ per year.

Location(s): [[mfield3

The role of a Site Reliability Engineer (SRE) at BAE Systems Digital Intelligence involves combining operations and software engineering to automate system support and enhance reliability for a key national security customer.

The SRE team works on continuous improvement of system health and performance, providing monitoring, automation, and incident response.

Responsibilities

  • Support and maintain essential services that support core mission applications, proactively improving availability, performance and stability.
  • Participate in a 24/7 on‑call rota, providing support for critical production systems outside of business hours, with additional on‑call allowances and overtime benefits.
  • Identify and automate repetitive tasks, developing tools to improve application health and reduce manual operation.
  • Collaborate with development teams to advise on good design practices for scalable, resilient systems.
  • Design and deploy monitoring products, creating custom tools as needed to provide comprehensive, intelligent observations that demonstrate daily improvements.
  • Participate in the wider Dev Ops/SRE community within the organization.

Qualifications

  • Experience in web development and object‑oriented programming.
  • Knowledge of database technologies such as Oracle SQL, Mongo DB, and Postgre SQL.
  • Proficiency with Linux and Windows command lines (Bash, Power Shell).
  • Experience monitoring large systems using Grafana, Prometheus, ELK, and Splunk.
  • Experience working in Agile teams and using Atlassian tooling.
  • Ability to diagnose and troubleshoot application issues causing service outages.
  • Strong troubleshooting skills across stack layers.
  • Understanding of ITIL processes.
  • Experience with micro‑services architectures, Docker and container platforms such as Open Shift and Kubernetes.
  • Awareness of emerging technology trends and curiosity to adopt cutting‑edge tools.
  • Security Clearance

Applicants must hold an active e DV clearance before applying.

EEO Statement

BAE Systems is an equal opportunity employer; we welcome applicants from all backgrounds.

#J-18808-Ljbffr

Site Reliability Engineer – NS London employer: 慨正橡扯

At 慨正橡扯, we pride ourselves on being an exceptional employer that champions innovation and collaboration in the field of Behavioral Economics and Retirement Research. Our hybrid working model not only offers flexibility but also nurtures a vibrant work culture where employees are encouraged to grow and develop their skills through meaningful projects and leadership opportunities. Join us in Europe, where your expertise will directly contribute to enhancing investor outcomes and shaping impactful business strategies.

Contact Details:

慨正橡扯 Recruitment Team

We think you need these skills to ace Site Reliability Engineer – NS London

Site Reliability Engineering
Operations and Software Engineering
System Automation
Monitoring and Incident Response
Web Development
Object-Oriented Programming
Database Technologies (Oracle SQL, MongoDB, PostgreSQL)