Senior Engineering Manager (Capacity Engineering)

Senior Engineering Manager (Capacity Engineering)

Full-Time 500 - 500 £ / month (est.) No working from home possible
A

At a Glance

  • Tasks: Lead a team to optimise and manage cutting-edge infrastructure systems.
  • Company: Join Anthropic, a leader in AI infrastructure management.
  • Benefits: Enjoy comprehensive health benefits, flexible time off, and competitive salary packages.
  • Other info: Dynamic work environment with opportunities for growth and collaboration.
  • Why this job: Make a real impact on innovative projects while developing your leadership skills.
  • Qualifications: Experience in managing engineering teams and a strong technical background required.

The predicted salary is between 500 - 500 £ per month.

  • Anthropic manages one of the largest and fastest-growing infrastructure fleets in the industry — spanning multiple accelerator families, CPU families, and clouds.

The Capacity Engineering team is responsible for making sure all of our infrastructure resources are accounted for, well-utilized, and efficiently allocated

  • We own the data, tooling, and operational systems that let Anthropic plan, measure, and maximize utilization across first-party and third-party compute — one of the company’s largest areas of spend
  • As the Senior Engineering Manager for Capacity Engineering, you will lead the team that builds and operates these production systems.

You’ll set technical direction, grow and develop a team of senior and staff-level engineers, and be accountable for the reliability and correctness of surfaces that leadership, research engineering, inference, infrastructure, and finance all depend on

  • This is a hands-on leadership role and will be onsite 5 days a week: we expect you to stay close enough to the systems to review designs, make sound architectural calls, and step into an incident when the team needs you — while spending most of your time on people, priorities, and cross-organizational alignment
  • The team’s work spans three overlapping areas, and you’ll be responsible for balancing investment across them as business priorities shift:
  • Data platform — Pipelines that ingest occupancy and utilization telemetry from Kubernetes clusters, normalize billing and usage across cloud providers, and serve the Big Query tables the rest of the org queries against.

Consumers range from research engineers to finance to leadership, so this is product work as much as engineering

  • Planning and Assurance — Making the state of the fleet legible and actionable in real time: cluster health tooling, capacity planning platforms, alerting on occupancy drops and allocation problems, and systemic fixes to scheduling and fragmentation
  • Efficiency — Measuring and improving how effectively every major workload uses the hardware it runs on, across training, inference, and evals.

Building benchmarking infrastructure and per-config baselines, then partnering with system-owning teams to close the gaps

  • Be hands-on, lead and grow the team.

Hire, onboard, coach, and retain senior and staff engineers.

Set clear expectations, give direct and timely feedback, run performance and leveling conversations, and build a team culture that values ownership, rigor, and collaboration

  • Champion your internal customers.

We build for our own use cases, so the teams that depend on our systems — research engineering, inference, infrastructure, and finance — are your customers.

Engage with them directly, bring what you learn back into the roadmap, and lead the team in building tools people genuinely want to use

  • Own the roadmap.

Translate company-level compute strategy into a prioritized engineering roadmap across data platform, planning and efficiency.

Make explicit trade-offs when priorities compete, and communicate them clearly upward and outward

  • Set the technical bar.

Review designs, weigh in on architecture, and hold the team to production standards — well-tested Python and SQL, latency and completeness SLOs, gap detection, and on-call that is sustainable

  • Run the team as a product organization.

Ensure the team gathers its own requirements, defines schema contracts, and designs for a wide range of consumers — from research engineers to a CFO.

  • Treat data quality and discoverability as first-class deliverables
  • Be the primary partner for cross-functional stakeholders.

Work closely with infrastructure, inference, research engineering, and finance leadership to align on capacity decisions, efficiency targets, and spend.

  • Represent the team’s data and recommendations to senior leadership
  • Drive operational excellence.

Own reliability and incident response for load-bearing systems, establish SLOs and on-call practices, and continuously reduce operational toil so the team can spend its time on higher-leverage work

  • Scale the function.

As the fleet diversifies (every new provider is a net-new integration), anticipate where the team needs to grow in headcount, skills, and systems — and make the case for it

Benefits

  • Comprehensive health, dental, and vision insurance for you and your dependents
  • Inclusive fertility benefits via Carrot Fertility
  • 22 weeks of paid parental leave
  • Flexible paid time off and absence policies
  • Mental health support for you and your dependents
  • Competitive salary and equity packages
  • Optional equity donation matching at a 1:1 ratio, up to 25% of your equity grant
  • Retirement plans with competitive matching
  • Life and income protection plans
  • $500/month flexible wellness and time saver stipend
  • Commuter benefits
  • Annual education stipend
  • Home office stipends
  • Relocation support for those moving for Anthropic
  • Daily meals and snacks in the office

Experience managing software or infrastructure engineering teams, including hiring senior engineers, managing performance, and developing people into larger scope A strong technical background in production systems — data engineering, infrastructure, distributed systems, or observability — with hands‑on experience you can still draw on when reviewing designs or debugging with the team Comfort owning operational responsibility for systems the company depends on, including on‑call and incident management A track record of setting and executing an engineering roadmap in an ambiguous, high‑autonomy environment with many stakeholders and shifting priorities Excellent communication skills: you can explain a utilization metric to a research engineer and a spend forecast to a CFO, and you can advocate clearly for your team’s priorities with senior leadership Familiarity with at least one major cloud provider (AWS, GCP, or Azure), Kubernetes‑based infrastructure, and modern observability stacks (e. g., Prometheus, Grafana)Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience We encourage you to apply even if you do not believe you meet every single qualification Experience leading teams working on capacity planning, resource management, product engineering or Fin Ops at a hyperscaler or in a large-scale ML environment Experience with multi-cloud billing and telemetry normalization (billing exports, reservation APIs, commitments, on‑demand capacity reservations)Familiarity with accelerator infrastructure — GPU metrics (DCGM), TPU utilization, or ML training and inference systems at the hardware level Background in scheduling, packing efficiency, or profiling‑driven optimization of large distributed workloads Experience building or leading internal data products with self‑service access, schema contracts, and documentation

#J-18808-Ljbffr

Senior Engineering Manager (Capacity Engineering) employer: Anthropic

At Anthropic, we pride ourselves on being an exceptional employer that fosters a culture of innovation and collaboration. Our team-oriented environment encourages personal growth and empowers employees to take ownership of their projects, making a meaningful impact in the tech landscape. Located in a vibrant area, we offer competitive benefits and unique opportunities for professional development, ensuring that our engineers thrive both personally and professionally.

A

Contact Details:

Anthropic Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Senior Engineering Manager (Capacity Engineering)

Join Local Tech Meetups

Get out there and mingle with fellow developers by joining local tech meetups. It’s a fantastic way to meet people who might be working at Anthropic or know someone who does. Plus, you can pick up some trendy tech skills and trends while you're at it!

Contribute to Open Source Projects

Show off your coding chops by jumping into open-source projects. Not only does this give you practical experience, but it also gets you noticed in the dev community. You'll create a killer portfolio that speaks volumes about your skills to Anthropic.

Tap into Online Developer Communities

Don’t underestimate the power of online developer communities like GitHub, Stack Overflow, and even Reddit. Participate in discussions, share your projects, and build your visibility. We can often find opportunities through these channels that can lead to a full-time gig at companies like Anthropic.

Explore Job Boards Specifically for Tech Roles

Keep your eyes peeled on job boards that focus on tech roles. Sites like TechCareers or Stack Overflow Jobs can often have listings for companies like Anthropic that might not show up on broader job sites. Make it a habit to check these regularly, and don’t hesitate to apply directly through our website!

We think you need these skills to ace Senior Engineering Manager (Capacity Engineering)

Team Leadership
Technical Direction
Production Systems Management
Data Engineering
Infrastructure Management
Distributed Systems
Observability

Some tips for your application 🫡

Show off your coding skills:When applying for a software engineering role, it's super important to showcase your coding skills. Make sure your CV includes your tech stack, any relevant programming languages you’re comfortable with, and examples of projects you've worked on. If you have a GitHub profile, link it up! We love to see code in action.

Tailor your portfolio:For a full-time role, we’d expect to see some solid examples of your work in your portfolio. Make sure to include at least two or three projects that highlight your problem-solving skills and your ability to work with different technologies. Focus on the projects that are most relevant to the position at Anthropic.

Craft a killer cover letter:Your cover letter is your chance to stand out—make it personal! Explain why you want to work at Anthropic and how your skills align with the role. Show us your passion for software development. We dig enthusiastic candidates who understand the value of collaboration and continuous learning!

Be clear and concise:When it comes to writing your CV and cover letter, clarity is key. Avoid jargon that could confuse us and stick to simple, direct language. Highlight your achievements with quantifiable results where possible, and keep everything easy to read. A well-organised application goes a long way!

How to prepare for a job interview at Anthropic

Brush Up on Your Coding Skills

For a full-time software engineering role, it's crucial that we stay sharp with our coding abilities. Expect technical questions that might involve solving problems on the spot or discussing algorithms. Practise on platforms like LeetCode or HackerRank to get comfortable with the types of questions that often come up.

Know Your Tools and Frameworks

Make sure we’re well-acquainted with the tools and technologies listed in the job description. Familiarise ourselves with any specific frameworks or programming languages mentioned. If Anthropic uses React or Node.js, for instance, be ready to discuss how we’ve used them in previous projects or coursework.

Showcase Your Projects

Bring along a portfolio that highlights our best work. This could be code samples, GitHub repositories, or any side projects we’ve built. Make sure we can talk through our thought process for each project, especially the challenges we faced and how we solved them—this shows our problem-solving skills in action.

Prepare for Behavioural Questions

While technical skills are key, full-time positions also require cultural fit. Be ready to discuss our previous experiences and how we handle teamwork, conflict, and deadlines. Brush up on the STAR method—Situation, Task, Action, Result—to clearly articulate our past experiences when discussing how we've contributed to a team.