Senior Production Engineer in London

Senior Production Engineer in London

London Full-Time 63000 - 77000 £ / year (est.) No working from home possible
C

At a Glance

  • Tasks: Ensure production systems are reliable and innovate solutions to enhance operational excellence.
  • Company: Join Clear Street, a cutting-edge fintech transforming market access for investors.
  • Benefits: Competitive salary, flexible work options, and opportunities for professional growth.
  • Other info: Collaborative team environment focused on building resilient systems.
  • Why this job: Make a real impact on production reliability while working with modern technologies.
  • Qualifications: Strong Python skills and experience in SRE or Production Engineering required.

The predicted salary is between 63000 - 77000 £ per year.

About Clear Street

Clear Street’s mission is to give every sophisticated investor access to every asset, in every market, through a unified platform built for speed, transparency and scale.

We give our clients the technology, tools, and service once reserved for the largest institutions, rebuilt with modern infrastructure.

Our single, cloud-native, end-to-end capital markets platform powers investor growth today and is transforming how they can interact with markets tomorrow.

For more information, visit https://clearstreet. io .

The Role

As a Production Engineer, you sit at the intersection of software reliability and operational excellence.

You own the health, resilience, and recovery of our production systems—while spending equal energy innovating solutions that eliminate human toil, reduce incident blast radius, and raise the reliability bar across the entire platform.

You will partner closely with engineering, operations, and business teams to understand daily pain points and translate them into lasting automated solutions.

Half your time is spent in the trenches—supporting production, responding to incidents, and deeply understanding how our
systems behave under real conditions.

The other half is yours to build: automation, tooling, and observability platforms that make tomorrow' s on-call shift meaningfully easier than today's.

You will work on challenges like

  • Design and build comprehensive monitoring and observability platforms that surface the right signal at the right time—eliminating alert fatigue and accelerating root-cause analysis.
  • Develop intelligent automation and self-healing capabilities that diagnose issues, trigger recovery workflows, and reduce mean time to recovery (MTTR) without manual intervention.
  • Analyze incidents, identify systemic trends, and engineer solutions that prevent entire classes of failures from recurring.
  • Build reusable runbooks, diagnostic tooling, and recovery playbooks that turn tribal knowledge into scalable platform capabilities.
  • Create golden-path operational workflows—making the safest, most reliable path also the easiest one for engineering teams to follow.
  • Partner with Platform Engineering to influence CI/CD pipelines, deployment safety, and infrastructure resilience from a production reliability perspective.
  • Champion Infrastructure as Code, Git Ops, and SRE best practices while helping teams adopt modern engineering workflows.
  • Continuously measure production health through SLIs/SLOs/SLAs, and drive engineering priorities based on reliability data.
  • Explore emerging technologies—including AI-assisted diagnostics and developer tooling—that transform how we operate production systems.
  • The Team

We believe resilient systems are built by engineers who understand them end to end.

Our Production Engineering team is the first and last line of defense for our production platform.

We treat reliability as a product, with uptime and engineer experience as our north stars.

We combine the discipline of SRE with a builder's mindset: when we see a recurring problem, we build a solution—not a workaround.

You will work across every engineering and operations team to understand failure modes, quantify reliability gaps, and build platform capabilities that scale with the organization.

Whether it's reducing MTTR from hours to minutes, building self-service diagnostic tools, or designing proactive alerting that catches issues before customers notice, your work will have immediate, measurable impact.

If you're passionate about making production systems invisible to end users—and you get energy from both firefighting and building the systems that make fires less likely—you'll thrive here.

What We're Looking For

We're looking for engineers who combine operational instinct with a builder's discipline.

You should have

  • Strong hands-on Python skills—this is your primary language for automation and tooling.
  • Experience in SRE, Production Engineering, Platform Engineering, or a related discipline with direct production ownership.
  • Proven track record of building automation and diagnostic tooling that improved recovery times or reduced operational toil.
  • Deep familiarity with cloud-native technologies—Kubernetes, containers, distributed systems—and how they fail in production.
  • Experience with observability platforms such as Datadog, and a strong intuition for what "good" monitoring looks like.
  • Exposure to Infrastructure as Code (Terraform) and Git Ops-based deployment workflows (Argo CD, Git Hub Actions, or similar).
  • Familiarity with the broader technology stack: Java, Go, Kafka, Redis, Snowflake, and Postgres.
  • Strong analytical and problem-solving skills—you thrive on ambiguous, high-stakes production problems.
  • A product mindset applied to operational tooling: you think about usability, adoption, and documentation when building internal solutions.
  • Excellent communication skills and the ability to work fluidly across engineering, operations, and business stakeholders.
  • Self-starter mentality—you identify opportunities, take initiative, and deliver with minimal supervision.
  • Curiosity and a continuous learning mindset; fintech or financial industry background is a plus.
  • The Technology You'll Work With
  • You'll operate and build on a modern cloud-native platform that includes:
  • Kubernetes & AWS
  • Terraform & Argo CD
  • Git Hub Actions
  • Kafka, Redis
  • Postgre SQL & Snowflake
  • Datadog
  • Python, Go, Java
  • g RPC & Protobuf
  • Internal Platform APIs and Developer Tooling
  • What Success Looks Like

Within your first year, you'll have made a measurable impact on production reliability. Success looks like:

  • Reducing mean time to detection (MTTD) and mean time to recovery (MTTR) across key production systems.
  • Building automation that handles a meaningful percentage of incident scenarios without human intervention.
  • Becoming a trusted subject matter expert for core platform components and their failure modes.
  • Delivering observability and diagnostic tools that other engineers actually use and depend on.
  • Establishing SLO baselines and driving engineering investment based on reliability data.
  • Spending less of your time—and your teammates time—on repetitive manual toil.

Your impact won't be measured by the number of incidents you respond to—it will be measured by how reliably our systems run and how quickly we recover when they don't.

Senior Production Engineer in London employer: Clear Street

Clear Street is an exceptional employer that fosters a culture of innovation and collaboration, empowering its employees to take ownership of their work while driving impactful solutions in the fast-paced fintech environment. With a strong focus on employee growth, Clear Street offers opportunities for continuous learning and development, alongside a commitment to work-life balance and a supportive team atmosphere. Located in a vibrant tech hub, employees benefit from access to cutting-edge technologies and a dynamic community that encourages creativity and excellence.

C

Contact Details:

Clear Street Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Senior Production Engineer in London

Join Local Tech Meetups

Get out there and mingle with fellow developers by joining local tech meetups. It’s a fantastic way to meet people who might be working at Clear Street or know someone who does. Plus, you can pick up some trendy tech skills and trends while you're at it!

Contribute to Open Source Projects

Show off your coding chops by jumping into open-source projects. Not only does this give you practical experience, but it also gets you noticed in the dev community. You'll create a killer portfolio that speaks volumes about your skills to Clear Street.

Tap into Online Developer Communities

Don’t underestimate the power of online developer communities like GitHub, Stack Overflow, and even Reddit. Participate in discussions, share your projects, and build your visibility. We can often find opportunities through these channels that can lead to a full-time gig at companies like Clear Street.

Explore Job Boards Specifically for Tech Roles

Keep your eyes peeled on job boards that focus on tech roles. Sites like TechCareers or Stack Overflow Jobs can often have listings for companies like Clear Street that might not show up on broader job sites. Make it a habit to check these regularly, and don’t hesitate to apply directly through our website!

We think you need these skills to ace Senior Production Engineer in London

Python
SRE
Production Engineering
Platform Engineering
Cloud-native Technologies
Kubernetes
Containers

Some tips for your application 🫡

Show off your coding skills:When applying for a software engineering role, it's super important to showcase your coding skills. Make sure your CV includes your tech stack, any relevant programming languages you’re comfortable with, and examples of projects you've worked on. If you have a GitHub profile, link it up! We love to see code in action.

Tailor your portfolio:For a full-time role, we’d expect to see some solid examples of your work in your portfolio. Make sure to include at least two or three projects that highlight your problem-solving skills and your ability to work with different technologies. Focus on the projects that are most relevant to the position at Clear Street.

Craft a killer cover letter:Your cover letter is your chance to stand out—make it personal! Explain why you want to work at Clear Street and how your skills align with the role. Show us your passion for software development. We dig enthusiastic candidates who understand the value of collaboration and continuous learning!

Be clear and concise:When it comes to writing your CV and cover letter, clarity is key. Avoid jargon that could confuse us and stick to simple, direct language. Highlight your achievements with quantifiable results where possible, and keep everything easy to read. A well-organised application goes a long way!

How to prepare for a job interview at Clear Street

Brush Up on Your Coding Skills

For a full-time software engineering role, it's crucial that we stay sharp with our coding abilities. Expect technical questions that might involve solving problems on the spot or discussing algorithms. Practise on platforms like LeetCode or HackerRank to get comfortable with the types of questions that often come up.

Know Your Tools and Frameworks

Make sure we’re well-acquainted with the tools and technologies listed in the job description. Familiarise ourselves with any specific frameworks or programming languages mentioned. If Clear Street uses React or Node.js, for instance, be ready to discuss how we’ve used them in previous projects or coursework.

Showcase Your Projects

Bring along a portfolio that highlights our best work. This could be code samples, GitHub repositories, or any side projects we’ve built. Make sure we can talk through our thought process for each project, especially the challenges we faced and how we solved them—this shows our problem-solving skills in action.

Prepare for Behavioural Questions

While technical skills are key, full-time positions also require cultural fit. Be ready to discuss our previous experiences and how we handle teamwork, conflict, and deadlines. Brush up on the STAR method—Situation, Task, Action, Result—to clearly articulate our past experiences when discussing how we've contributed to a team.