At a Glance
- Tasks: Resolve escalations and incidents while ensuring top-notch customer satisfaction.
- Company: Join a leading tech firm focused on Azure operations and support.
- Benefits: Attractive salary, flexible working options, and opportunities for professional growth.
- Other info: Dynamic role with mentorship opportunities and a focus on innovation.
- Why this job: Be at the forefront of Azure technology and make a real difference in customer experiences.
- Qualifications: 7-10 years IT experience with strong Azure operations and troubleshooting skills.
The predicted salary is between 60000 - 80000 Β£ per year.
To ensure on-time resolution of escalations/incidents through efficient analysis as per the SLA and quality norms and ensure positive customer satisfaction.
Key Responsibilities
- Common Operations Responsibilities (Applicable to All Roles β L2 & L3)
- OS & Platform Patching: Monthly/weekly OS and platform patching of Azure VMs and PaaS services. Patch validation, compliance tracking, and post-patch verification.
- Backup & Restore Operations: Daily monitoring of Azure Backup jobs (VMs, PaaS-supported workloads). Backup validation, restore testing, and failure remediation.
- Monitoring & Alert Management: Proactive monitoring using Azure Monitor, Log Analytics, Alerts. Availability, performance, and capacity monitoring. Alert tuning and noise reduction.
- Basic Automation: PowerShell / Azure CLI / Python scripts for routine operational tasks. Automation of health checks, reports, and repetitive activities.
- Weekly / Monthly Data Management: Capacity utilization reports (compute, storage, backup). Monthly operational dashboards and health reports.
- Incident & Troubleshooting: Troubleshooting Azure compute, network, storage, and PaaS services. Support for AVS and Azure Stack HCI platforms.
- Knowledge Management: Create and maintain KB articles, SOPs, and runbooks. Update documentation after incidents and changes.
- Problem Management: Identify recurring issues and raise problem records. Support root cause identification and preventive fixes.
- Priority Incident Management: End-to-end ownership of P1 / P2 incidents. Incident bridge participation and customer coordination.
- Customer Communication: Handle customer calls, MSR / Major Service Restoration. Provide timely status updates during incidents.
- Operational Governance: Participate in daily, weekly, bi-weekly, and monthly operational calls. QBR, capacity planning, and service review meetings.
- Azure Administrator β L3 (IaaS / PaaS | Operations & Escalation)
- Role Summary: L3 Azure Administrator responsible for complex escalations, RCA, automation, security, and platform stability.
- Key Responsibilities: L3 ownership of complex Azure incidents and escalations. Deep troubleshooting of Azure IaaS, PaaS and DevOps. Patch strategy planning and compliance reporting. Backup architecture review, restore validation, and DR drills. Advanced automation and monitoring improvements. Problem management and preventive action implementation. Lead MSR, QBR, and capacity planning discussions. Mentor L2 teams and improve SOPs / KB articles.
- Skills & Experience: 7β10+ years overall IT experience, 4β6+ years Azure operations expertise, strong RCA, troubleshooting, and security skills, advanced PowerShell / Azure CLI / Python automation, Azure networking and governance expertise.
Subject Matter Expert (Support&Ops) in London employer: HCL Technologies Limited
As a Track Lead in MDM Intune, you will join a forward-thinking organisation that prioritises employee growth and development, offering access to higher education programmes and skills resources. Our collaborative work culture fosters innovation and ensures that your contributions directly enhance the stability and security of our endpoint ecosystem, impacting nearly all users across the enterprise. With generous personal time off and family benefits, we are committed to supporting a balanced and fulfilling work-life experience.