At a Glance
- Tasks: Lead the development of our observability platform and enhance fleet management.
- Company: Join Nscale, a pioneering GenAI cloud platform with a culture of innovation.
- Benefits: Inclusive workplace, competitive salary, and opportunities for professional growth.
- Other info: Diverse team environment with a commitment to equity and inclusion.
- Why this job: Make a real impact in AI technology and work with cutting-edge tools.
- Qualifications: 5-8 years in product management with experience in observability and infrastructure.
The predicted salary is between 66150 - 80850 £ per year.
About Nscale
Nscale is taking on the hyperscalers by building a vertically integrated Gen AI cloud platform.
We own the data centres, software, and applications that power today's AI stack using sustainable technology solutions.
We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency.
As an Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work.
Collaboration is key, and we work together swiftly and respectfully, embracing adaptability and resilience in all we do.
About the role
Technical Product Managers at Nscale own the definition, delivery, and ongoing evolution of a slice of the Nscale platform, partnering with engineering, design, and go-to-market to turn customer and operational problems into shippable outcomes.
As a Senior Technical Product Manager for Observability, you own the platform that gives customers and internal operators real-time visibility into their GPU fleet: the telemetry pipeline that scrapes data from physical infrastructure, the aggregation and storage layer, and the observability surfaces (logs, metrics, and traces) that enable fleet management, incident response, and alerting at scale.
You partner daily with Fleet Software, Network Engineering, Data Centre Operations, and customer teams to make fleet health visible, actionable, and reliable as Nscale scales from a handful of deployments to a globally distributed fleet.
- What you'll be doing
- Own the roadmap for Nscale's observability platform: the telemetry pipeline, log and metrics aggregation, trace collection, and customer facing APIs and dashboards that surface fleet health to customers and operators.
- Define how logs, metrics, and traces are captured from physical infrastructure, aggregated, and surfaced through the observability platform to enable customers to manage their fleet and handle incidents.
- Own alerting strategy and optimisation: define what matters, reduce noise, and ensure the right signal reaches the right person at the right time.
- Capture and prioritise new telemetry requirements as the fleet scales, working with engineering to extend coverage across new hardware, sites, and deployment types.
- Shadow incident reviews and site operations to turn recurring manual effort and visibility gaps into platform capabilities.
- Define and drive the metrics that matter: alert signal-to-noise ratio, time-to-detect, time-to-resolve, telemetry coverage, and platform reliability.
- Mentor junior PMs and raise the bar for PRDs, reviews, and product decisions across the team.
- What you need
- 5-8 years in product management, with a track record owning significant areas in observability, infrastructure, or operations-facing products.
- Demonstrated experience building observability stacks: you have owned a product that captures and surfaces logs, metrics, and traces at scale, and you understand the architectural and UX tradeoffs involved.
- Hands-on experience with Prometheus, Loki, Mimir, Datadog, Grafana, or Open Telemetry.
- Experience with deployment tooling in a data centre or infrastructure context, including provisioning workflows, networking automation, or zero-touch deployment pipelines.
- Experience building for operators and delivery teams (design engineers, project controllers, PMs, SREs, DC technicians) and a genuine appetite for their workflows.
- Strong technical fluency: you can lead architecture and trade-off discussions across telemetry pipelines, time-series storage, alerting systems, and observability integrations.
- A record of moving ambiguous operational problems to shipped outcomes that measurably improve visibility, incident response, or fleet reliability.
- Excellent written and verbal communication across engineers, operators, and executives.
- Nice to haves
- Broader observability problem domain experience across different toolsets beyond the above stack.
- Familiarity with bare-metal provisioning tools (Open Stack Ironic, MAAS, or similar) or network automation tooling (Net Box, Nautobot, or similar).
Degree in CS or engineering, or prior experience as an engineer, SRE, or infrastructure operator.
- Familiarity with GPU or accelerated compute infrastructure, data centre operations, or hyperscaler-style deployment at scale.
- ITSM: Jira Service Management, Service Now, Zendesk, or Freshservice.
- Experience in high-growth environments where the product is being built alongside the fleet it monitors.
Join Nscale as we build a world-class AI cloud platform. If you're excited about owning the software that turns contracts into live GPU capacity, we'd love to hear from you!
At Nscale, we are committed to fostering an inclusive, diverse, and equitable workplace.
We believe that a variety of perspectives enriches our work environment, and we encourage applications from candidates of all backgrounds, experiences, and abilities.
We strongly encourage applications from people of colour, the LGBTQ+ community, people with disabilities, neurodivergent people, parents, carers, and people from lower socio-economic backgrounds.
If there’s anything we can do to accommodate your specific situation, please let us know.
The responsibilities outlined in this job description are not exhaustive and are intended to provide a general overview of the position.
The employee may be required to perform additional duties, tasks, and responsibilities as assigned by management, consistent with the skills and qualifications required for the role.
For information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice: Here.
#J-18808-Ljbffr
Senior Product Manager, Observability employer: AI Startups UK
Lovable is an exceptional employer for data scientists, particularly in our dynamic London growth team. With a culture of extreme ownership and low-ego collaboration, we empower our employees to drive impactful growth metrics while working alongside talented engineers. Our commitment to innovation and rapid experimentation offers unparalleled opportunities for professional development, making Lovable a truly rewarding place to build your career.
StudySmarter Expert Advice🤫
We think this is how you could land Senior Product Manager, Observability
✨Join Local Tech Meetups
Get out there and mingle with fellow developers by joining local tech meetups. It’s a fantastic way to meet people who might be working at AI Startups UK or know someone who does. Plus, you can pick up some trendy tech skills and trends while you're at it!
✨Contribute to Open Source Projects
Show off your coding chops by jumping into open-source projects. Not only does this give you practical experience, but it also gets you noticed in the dev community. You'll create a killer portfolio that speaks volumes about your skills to AI Startups UK.
✨Tap into Online Developer Communities
Don’t underestimate the power of online developer communities like GitHub, Stack Overflow, and even Reddit. Participate in discussions, share your projects, and build your visibility. We can often find opportunities through these channels that can lead to a full-time gig at companies like AI Startups UK.
✨Explore Job Boards Specifically for Tech Roles
Keep your eyes peeled on job boards that focus on tech roles. Sites like TechCareers or Stack Overflow Jobs can often have listings for companies like AI Startups UK that might not show up on broader job sites. Make it a habit to check these regularly, and don’t hesitate to apply directly through our website!
We think you need these skills to ace Senior Product Manager, Observability
Some tips for your application 🫡
Show off your coding skills:When applying for a software engineering role, it's super important to showcase your coding skills. Make sure your CV includes your tech stack, any relevant programming languages you’re comfortable with, and examples of projects you've worked on. If you have a GitHub profile, link it up! We love to see code in action.
Tailor your portfolio:For a full-time role, we’d expect to see some solid examples of your work in your portfolio. Make sure to include at least two or three projects that highlight your problem-solving skills and your ability to work with different technologies. Focus on the projects that are most relevant to the position at AI Startups UK.
Craft a killer cover letter:Your cover letter is your chance to stand out—make it personal! Explain why you want to work at AI Startups UK and how your skills align with the role. Show us your passion for software development. We dig enthusiastic candidates who understand the value of collaboration and continuous learning!
Be clear and concise:When it comes to writing your CV and cover letter, clarity is key. Avoid jargon that could confuse us and stick to simple, direct language. Highlight your achievements with quantifiable results where possible, and keep everything easy to read. A well-organised application goes a long way!
How to prepare for a job interview at AI Startups UK
✨Brush Up on Your Coding Skills
For a full-time software engineering role, it's crucial that we stay sharp with our coding abilities. Expect technical questions that might involve solving problems on the spot or discussing algorithms. Practise on platforms like LeetCode or HackerRank to get comfortable with the types of questions that often come up.
✨Know Your Tools and Frameworks
Make sure we’re well-acquainted with the tools and technologies listed in the job description. Familiarise ourselves with any specific frameworks or programming languages mentioned. If AI Startups UK uses React or Node.js, for instance, be ready to discuss how we’ve used them in previous projects or coursework.
✨Showcase Your Projects
Bring along a portfolio that highlights our best work. This could be code samples, GitHub repositories, or any side projects we’ve built. Make sure we can talk through our thought process for each project, especially the challenges we faced and how we solved them—this shows our problem-solving skills in action.
✨Prepare for Behavioural Questions
While technical skills are key, full-time positions also require cultural fit. Be ready to discuss our previous experiences and how we handle teamwork, conflict, and deadlines. Brush up on the STAR method—Situation, Task, Action, Result—to clearly articulate our past experiences when discussing how we've contributed to a team.