Cloud Infrastructure & Site Reliability Engineer
Contract | Sheffield | Hybrid | Β£400+ p/d
We are hiring a Cloud Infrastructure & Site Reliability Engineer to support and enhance a large-scale cybersecurity data lake and analytics platform. This role sits within a global Platform & Data Engineering team and focuses on cloud infrastructure, platform engineering, and operational reliability in Azure.
The position is hands-on and suited to engineers with strong cloud foundations, experience supporting production environments, and the ability to work across data, infrastructure, and platform operations.
Key Responsibilities
- Operate, maintain, and improve an enterprise-scale Azure cybersecurity data platform.
- Build and manage Azure landing zones/workspaces for secure application deployment.
- Integrate and configure cloud services including IAM, monitoring, OS, containers.
- Support CI/CD pipelines and continuous testing activities.
- Provide day-to-day troubleshooting across the Azure tenant and platform.
- Ensure platform reliability, resiliency, compliance, and operational risk alignment.
- Maintain documentation, runbooks, and operational artefacts.
- Collaborate with cloud service teams and internal stakeholders to gather requirements.
Skills & Experience
- Experience with SRE principles and Azure DevOps.
- Linux administration (preferably Red Hat).
- Understanding of OS internals, storage, CPU architectures, and network fundamentals.
- Hands-on experience building and supporting Azure infrastructure at scale.
- Familiarity with high-availability web infrastructure (proxies, reverse proxies).
- Knowledge of data governance, quality, and controls.
- Experience with IAM, cloud compliance, monitoring, and auditing tools.
- Strong troubleshooting skills across cloud, services, and application stacks.
- Exposure to Azure Data Factory, Databricks, Functions, Kubernetes Service, Logic Apps, Monitor, Log Analytics, Synapse, PowerBI.
Background
- Degree in Computer Science, Software Engineering, Data Science, or STEM preferred.
- Candidates without a degree will be considered with relevant experience or certifications.
- Experience in regulated industries (e.g., Financial Services) beneficial.
#J-18808-Ljbffr
Site Reliability Engineer (Sheffield) employer: Caspian One
Join a pioneering engineering team in Sheffield, where you'll have the unique opportunity to shape the internal developer experience for one of the world's largest financial institutions. With a strong focus on innovation and continuous improvement, our culture promotes autonomy and collaboration, allowing you to influence major strategic programmes while working with cutting-edge technologies. Enjoy a hybrid working model that balances flexibility with regular office engagement, ensuring a rewarding and impactful career path.