At a Glance
- Tasks: Design and manage OpenSearch clusters, enhance monitoring solutions, and collaborate with engineering teams.
- Company: FDM, a global leader in tech and business talent solutions.
- Benefits: Career coaching, global assignments, upskilling opportunities, annual leave, and pension.
- Other info: Diverse and inclusive workplace with excellent career growth potential.
- Why this job: Join a dynamic team to drive observability improvements in the finance sector.
- Qualifications: Experience with OpenSearch, Grafana, and strong scripting skills required.
The predicted salary is between 63000 - 77000 £ per year.
About The Role
FDM is a global business and technology consultancy seeking a Site Reliability Engineer to work for our client within the Finance sector.
This is initially a 6 month contract with very good prospects to extend and will be a hybrid role that will be based in London.
Our client is seeking an experienced Site Reliability Engineer (SRE) with a strong focus on Observability and Monitoring Platforms.
The successful candidate will play a key role in enhancing the organisation's monitoring, alerting, and operational visibility capabilities across critical engineering systems.
This role requires hands-on expertise in the deployment, administration, and optimisation of Open Search, alongside experience with Grafana, Geneos, and automation/scripting technologies.
Particular emphasis will be placed on the candidate's ability to design, deploy, and support enterprise-grade Open Search environments.
Responsibilities
- Lead the design, deployment, configuration, and ongoing management of Open Search clusters and associated observability tooling.
- Develop and maintain scalable monitoring, logging, and alerting solutions for business-critical applications and infrastructure.
- Build and enhance observability dashboards using Grafana.
- Support and optimise existing Geneos monitoring implementations.
- Create and maintain automation scripts to streamline operational processes and improve reliability.
- Collaborate with engineering, infrastructure, and support teams to improve system resilience and operational performance.
- Define and implement SRE best practices, including monitoring standards, alert management, incident response, and operational readiness.
- Perform troubleshooting and root cause analysis of platform and application issues.
- Support capacity planning, performance tuning, and platform optimisation initiatives.
- Contribute to documentation, operational procedures, and knowledge sharing within the engineering team.
About You
- Open Search (Critical Requirement)
- Extensive hands-on experience deploying and managing Open Search in production environments.
- Deep understanding of Open Search architecture, cluster design, indexing strategies, shard management, and performance tuning.
- Experience implementing log aggregation, search, analytics, and observability use cases using Open Search.
- Knowledge of Open Search security, access controls, backups, upgrades, and operational best practices.
- Monitoring & Observability
- Strong experience with Grafana, including dashboard development, alerting, and data source integration.
- Experience with enterprise monitoring platforms, specifically Geneos.
- Understanding of modern observability principles, including metrics, logs, traces, alerting, and service health monitoring.
- Scripting & Automation
• Strong scripting skills in one or more of
- Python
- Shell/Bash
- Power Shell
- Experience automating operational tasks and monitoring workflows.
- SRE / Platform Engineering
- Proven experience in an SRE, Platform Engineering, Dev Ops, or Infrastructure Engineering role.
- Strong troubleshooting and problem-solving capabilities.
- Experience supporting highly available and business-critical systems.
- Understanding of incident management, resilience engineering, and operational excellence practices.
Desirable Skills
- Experience with cloud platforms (Azure, AWS, or GCP).
- Knowledge of containerisation technologies (Docker, Kubernetes).
- Experience with CI/CD pipelines and Infrastructure as Code.
- Experience working within financial services or regulated environments.
- Familiarity with Elasticsearch ecosystems and migration strategies to Open Search.
- Candidate Profile
The ideal candidate will be a hands-on engineer who combines deep technical expertise with a pragmatic operational mindset.
They will be comfortable working independently, driving observability improvements, and collaborating across engineering teams to deliver reliable and scalable monitoring solutions.
- Key attributes
- Strong ownership mentality.
- Excellent analytical and troubleshooting skills.
- Effective stakeholder communication.
- Ability to operate in fast-paced production environments.
- Focus on reliability, automation, and continuous improvements
About Us
FDM is an award-winning global leader in tech and business talent solutions, backed by more than 35 years of industry experience.
We have centres across Europe, North America, and Asia-Pacific, and a global workforce of over 2500 employees.
FDM has shown exponential growth throughout the years, firmly establishing itself as an award-winning employer, currently listed on the FTSE4Good Index and as a 2026 Financial Times UK ‘Best Employer’.
Diversity and Inclusion the company is an equal opportunity employer, and all qualified applicants will receive consideration for employment without regard to race, colour, religion, sex, sexual orientation, national origin, age, disability, veteran status or any other status protected by federal, provincial or local laws.
- Why join us
- Career coaching, mentoring and access to upskilling throughout your entire FDM career
- Assignments with global companies and opportunities to work abroad
- Opportunity to re-skill and up-skill into new areas, develop non-linear career paths and build a skillset within your field
- Annual leave and work-place pension
- #J-18808-Ljbffr
Site Reliability Engineer- London employer: United States Digital Space LLC
United States Digital Space LLC is an exceptional employer, offering a dynamic work culture that prioritises innovation and collaboration in the heart of Greater London. With a strong focus on employee well-being and flexible work options, we provide ample opportunities for professional growth and development, making it an ideal environment for those looking to make a meaningful impact in the field of AI-enabled SaaS engineering.
Contact Details:
United States Digital Space LLC Recruitment Team
StudySmarter Expert Advice🤫
We think this is how you could land Site Reliability Engineer- London
✨Get Involved in Open-Source Projects
Diving into open-source projects is a brilliant way to showcase your skills and connect with other developers in the community. Not only will you beef up your GitHub profile but you might also catch the eye of someone at United States Digital Space LLC who values hands-on experience over just theory.
✨Attend Local Tech Meetups
Tech meetups are gold mines for networking and discovering job opportunities, especially in the fast-paced world of software engineering. Check out local listings for events in your area and don’t shy away from introducing yourself. This could lead directly to a temporary position at United States Digital Space LLC!
✨Showcase Your Work Online
With temporary roles, it’s all about standing out in a short space of time. Create a portfolio website where you highlight your projects and skills. Talk about your code, and provide links to your GitHub repositories. This will not only demonstrate your abilities but will also make it easier for recruiters at United States Digital Space LLC to see what you bring to the table.
✨Leverage Temporary Job Boards
Don’t forget to check specialised job boards for temporary software development roles. Websites like We Work Remotely or Remote OK often list short-term gigs that can be a perfect fit. Apply directly through our website as well, making sure your application is sharp—temporary roles can move fast!
We think you need these skills to ace Site Reliability Engineer- London
Some tips for your application 🫡
Show Off Your Tech Skills:Make sure your CV highlights your tech stack and any programming languages you’re proficient in. Include specifics about any frameworks or technologies you’ve worked with; they can make you stand out in the sea of applicants. It’s all about showing that you have the chops we need at United States Digital Space LLC!
Portfolio 2.0:Since you’re applying for a temporary gig, it’s super important to showcase a portfolio that highlights your best projects. Include links to GitHub or any personal projects that demonstrate what you can do in a real-world environment. This gives us a taste of your style and your problem-solving approach!
Keep It Brief and Relevant:With a temporary position, we want to see your ability to hit the ground running. Be concise in your CV and cover letter; stick to experiences that directly relate to the role. Highlight any previous temporary roles or freelance gigs that show your adaptability and quick learning!
Tailor Your Cover Letter:Don’t just send a generic cover letter. Personalise it for Site Reliability Engineer- London at United States Digital Space LLC! Mention why this temporary role excites you and how you see yourself contributing in the short run. Show us what you've got and why you're the one for this quick turn-around!
How to prepare for a job interview at United States Digital Space LLC
✨Nail the Technical Skills
For a software engineering role, you'll likely face technical questions or coding tasks during your interview. Brush up on the relevant programming languages and frameworks that United States Digital Space LLC uses, and don’t forget to practice some coding challenges on platforms like LeetCode or HackerRank. Showing your coding prowess can really make you stand out!
✨Prepare for System Design Questions
Even for a temporary role, having a grasp of system design principles can be crucial. Be ready to discuss how you would architect a software solution, including discussing trade-offs, scalability, and performance considerations. Having examples from previous projects can really show off your analytical thinking.
✨Demonstrate Your Adaptability
Since this is a temporary role, you'll want to emphasise your ability to hit the ground running. Highlight experiences where you quickly adapted to new technologies or teams. Let’s make it clear to the interviewers at United States Digital Space LLC that you can learn on the job and deliver results in a short timeframe!
✨Show Off Your Portfolio
Make sure to have a portfolio or GitHub ready showcasing your projects. Having tangible evidence of what you've done—be it personal projects, contributions to open-source, or previous work—can convey how capable you are. Tailor this for what might interest United States Digital Space LLC, so it's relevant and sparks conversation during your interview.