At a Glance
- Tasks: Lead the strategy and implementation of enterprise tooling during AWS migrations.
- Company: Join a forward-thinking tech company in London.
- Benefits: Competitive pay, flexible working arrangements, and opportunities for professional growth.
- Other info: Collaborative environment with strong leadership and career advancement potential.
- Why this job: Make a significant impact on cloud migration projects and enhance operational visibility.
- Qualifications: 14+ years in observability tooling and AWS architecture required.
Hope this email finds you well.
Please find attached the job description for 4 urgent roles. I need your help in sourcing sub-contractors for these roles based in London.
Kindly help on priority. Please keep these profiles preferably within IR35.
|
Role |
JR |
Start date |
Duration |
Rate |
Pay rate |
|
Lead for Enterprise tooling & observability for migration |
32945 |
26-Feb |
4-6 weeks |
600 -650 |
550 |
Job Description: Lead – Enterprise Tooling & Observability (14+ Years Experience)
Role Overview
We are looking for an experienced Enterprise Tooling & Observability Lead to drive the strategy, design, implementation, and modernization of enterprise monitoring, logging, APM, and operational tooling during and after large‑scale on‑prem to AWS cloud migrations.
The ideal candidate brings deep expertise across observability platforms, infrastructure/application monitoring, cloud‑native operations, and integration of enterprise tools into cloud architectures. This role ensures a seamless migration of tooling capabilities, enhanced visibility, and improved reliability in the AWS operating model.
Key Responsibilities
1. Tooling & Observability Strategy Across Migration Lifecycle
- Define the end-to-end tooling & observability architecture that supports pre‑migration, migration, and post‑migration operations.
- Assess current on‑prem tooling (monitoring, logging, APM, ITSM, event management) and define cloud‑aligned target tooling architecture for AWS.
- Build a unified observability roadmap covering metrics, logs, traces, dashboards, SLO/SLA monitoring, and event correlation.
2. Migration‑Aware Observability Design
- Identify tooling gaps that may arise during the migration of applications, networks, storage, and infrastructure.
- Ensure instrumentation readiness for applications moving via lift & shift, replatforming, containerization, and modernization.
- Define observability patterns for hybrid connectivity, multi-account AWS environments, and multi-region workloads.
3. AWS Cloud-Native Observability Integration
- Design and implement observability using AWS-native capabilities such as:
- CloudWatch, CloudTrail, X-Ray, VPC Flow Logs, GuardDuty, Security Hub
- Integration with AWS Control Tower/Organizations for enterprise-wide visibility
- Ensure seamless integration with third‑party enterprise tools such as:
- Datadog, Dynatrace, AppDynamics
- Splunk/ELK
- Prometheus/Grafana
- ServiceNow, Jira, PagerDuty
- Drive modernization of legacy monitoring solutions to cloud-native ecosystems.
4. Tooling Consolidation & Optimization
- Evaluate existing tooling footprint and identify opportunities for consolidation, cost reduction, and simplification.
- Standardize tooling patterns and create reusable templates/playbooks for AWS workloads.
- Drive automation for alerting, dashboards, health checks, and operational insights.
5. Reliability, Performance & SRE Alignment
- Collaborate with platform and SRE teams to enhance observability maturity (SLIs, SLOs, error budgets).
- Build proactive monitoring capabilities to reduce incidents, improve MTTR, and support predictive operations.
- Ensure the observability platform aligns with enterprise DR, HA, and performance engineering strategies.
6. Governance, Security & Compliance
- Ensure observability tooling adheres to enterprise security, compliance, data governance, and access control policies.
- Define audit‑ready logging strategies and ensure end‑to‑end traceability across hybrid and cloud environments.
- Build governance models for event noise reduction, alert hygiene, and service mapping accuracy.
7. Leadership & Stakeholder Management
- Lead cross-functional teams across application, infrastructure, cloud, DevOps, and security functions.
- Serve as the primary SME for observability decisions, guiding teams through architectural design and implementation.
- Present observability strategy, migration readiness, platform health, and maturity improvements to senior leadership.
- Mentor engineers and drive capability uplift across the organization.
Required Skills & Experience
Technical Expertise
- 14+ years of experience in enterprise monitoring, logging, APM, and observability tooling.
- Strong understanding of AWS architecture, cloud‑native monitoring tools, and hybrid observability.
- Experience with:
- APM platforms: Dynatrace, AppDynamics, Datadog
- Logging platforms: Splunk, ELK/Opensearch, CloudWatch Logs
- Metrics & telemetry: Prometheus, Grafana, OpenTelemetry
- Event management: ServiceNow, PagerDuty, Moogsoft, BigPanda
- Strong knowledge of instrumentation for distributed systems, microservices, containers (EKS, ECS), serverless workloads, and legacy systems.
Migration & Architecture Skills
- Proven experience supporting large-scale on‑prem to AWS migrations.
- Deep understanding of migration patterns and observability dependencies.
- Hands-on experience designing observability for multi-account AWS landing zones and multi-region architectures.
Soft Skills & Leadership
- Excellent communication, architectural documentation, and executive presentation skills.
- Ability to influence stakeholders across engineering, cloud, SRE, operations, and leadership.
- Experience leading cross-functional teams and managing vendor/tooling relationships.
Preferred Qualifications
- AWS Certified Solutions Architect / Cloud Practitioner / DevOps Engineer
- Certifications in observability platforms (Datadog, Dynatrace, Splunk, etc.)
- Knowledge of ITIL, SRE principles, and enterprise operational frameworks
- Experience with automation using Python, Terraform, CloudFormation (nice-to-have)
Success Indicators
- Smooth transition of observability and tooling through all migration waves.
- Enhanced end‑to‑end visibility across applications, networks, and infrastructure post‑migration.
- Reduction in incidents, MTTR, and monitoring gaps after migration to AWS.
- Standardized tooling practices aligned with enterprise governance and cloud architecture.
- Strong stakeholder confidence and measurable uplift in observability maturity.
Lead for Enterprise tooling & observability for migration in London employer: American IT Systems
Palantir is an exceptional employer, offering a dynamic work culture that fosters innovation and collaboration among high-complexity teams. With a focus on employee growth, we provide opportunities for professional development in cutting-edge technologies like Kubernetes and Terraform, all while contributing to meaningful projects that support UK Government programs in defense. Our London location not only offers a vibrant city life but also a unique chance to work on secure and scalable infrastructure that makes a real impact.
StudySmarter Expert Advice🤫
We think this is how you could land Lead for Enterprise tooling & observability for migration in London
✨Get Involved in Open-Source Projects
Diving into open-source projects is a brilliant way to showcase your skills and connect with other developers in the community. Not only will you beef up your GitHub profile but you might also catch the eye of someone at American IT Systems who values hands-on experience over just theory.
✨Attend Local Tech Meetups
Tech meetups are gold mines for networking and discovering job opportunities, especially in the fast-paced world of software engineering. Check out local listings for events in your area and don’t shy away from introducing yourself. This could lead directly to a temporary position at American IT Systems!
✨Showcase Your Work Online
With temporary roles, it’s all about standing out in a short space of time. Create a portfolio website where you highlight your projects and skills. Talk about your code, and provide links to your GitHub repositories. This will not only demonstrate your abilities but will also make it easier for recruiters at American IT Systems to see what you bring to the table.
✨Leverage Temporary Job Boards
Don’t forget to check specialised job boards for temporary software development roles. Websites like We Work Remotely or Remote OK often list short-term gigs that can be a perfect fit. Apply directly through our website as well, making sure your application is sharp—temporary roles can move fast!
We think you need these skills to ace Lead for Enterprise tooling & observability for migration in London
Some tips for your application 🫡
Show Off Your Tech Skills:Make sure your CV highlights your tech stack and any programming languages you’re proficient in. Include specifics about any frameworks or technologies you’ve worked with; they can make you stand out in the sea of applicants. It’s all about showing that you have the chops we need at American IT Systems!
Portfolio 2.0:Since you’re applying for a temporary gig, it’s super important to showcase a portfolio that highlights your best projects. Include links to GitHub or any personal projects that demonstrate what you can do in a real-world environment. This gives us a taste of your style and your problem-solving approach!
Keep It Brief and Relevant:With a temporary position, we want to see your ability to hit the ground running. Be concise in your CV and cover letter; stick to experiences that directly relate to the role. Highlight any previous temporary roles or freelance gigs that show your adaptability and quick learning!
Tailor Your Cover Letter:Don’t just send a generic cover letter. Personalise it for Lead for Enterprise tooling & observability for migration at American IT Systems! Mention why this temporary role excites you and how you see yourself contributing in the short run. Show us what you've got and why you're the one for this quick turn-around!
How to prepare for a job interview at American IT Systems
✨Nail the Technical Skills
For a software engineering role, you'll likely face technical questions or coding tasks during your interview. Brush up on the relevant programming languages and frameworks that American IT Systems uses, and don’t forget to practice some coding challenges on platforms like LeetCode or HackerRank. Showing your coding prowess can really make you stand out!
✨Prepare for System Design Questions
Even for a temporary role, having a grasp of system design principles can be crucial. Be ready to discuss how you would architect a software solution, including discussing trade-offs, scalability, and performance considerations. Having examples from previous projects can really show off your analytical thinking.
✨Demonstrate Your Adaptability
Since this is a temporary role, you'll want to emphasise your ability to hit the ground running. Highlight experiences where you quickly adapted to new technologies or teams. Let’s make it clear to the interviewers at American IT Systems that you can learn on the job and deliver results in a short timeframe!
✨Show Off Your Portfolio
Make sure to have a portfolio or GitHub ready showcasing your projects. Having tangible evidence of what you've done—be it personal projects, contributions to open-source, or previous work—can convey how capable you are. Tailor this for what might interest American IT Systems, so it's relevant and sparks conversation during your interview.