Cloud SRE/ Site Reliability Engineer (Azure Platform Engineering)
London | HybridWe're partnering with a global organisation investing heavily in cloud transformation, platform engineering and next-generation technology capabilities.
As part of a growing Cloud & Infrastructure team, we're looking for an experienced
Site Reliability Engineer to help design, build and evolve enterprise-scale Azure platforms that support critical business services across a global environment.
This is
not a traditional operations-focused SRE role. We're particularly interested in candidates with strong experience across
Azure infrastructure, networking, API Management, automation and platform engineering. The ideal candidate will enjoy building modern cloud platforms, driving automation and enabling emerging technologies, including AI-driven services.
What You'll Be Doing- Designing and supporting highly available, scalable Azure platforms
- Driving platform reliability, resilience and operational excellence
- Building Infrastructure as Code solutions using Terraform
- Developing and improving Azure DevOps CI/CD pipelines
- Implementing monitoring, observability and proactive alerting capabilities
- Supporting Azure networking, security and API Management services
- Designing disaster recovery, failover and business continuity solutions
- Reducing operational toil through automation and self-healing capabilities
- Working closely with Architecture, Security and Engineering teams to define platform standards and best practices
Key Technologies- Microsoft Azure
- Terraform
- Azure DevOps
- Azure Networking
- Azure API Management
- Azure Monitor / Log Analytics
- Azure Security & Identity
- AKS
- PowerShell, Python or Bash
What We're Looking For- Significant experience in a senior SRE, Platform Engineer, Cloud Engineer or Azure Infrastructure role
- Strong Azure architecture and engineering expertise
- Hands-on Terraform and Infrastructure as Code experience
- Experience building and automating cloud platforms at scale
- Strong understanding of networking, connectivity and cloud security
- Experience with observability, monitoring and incident management
- Knowledge of SRE principles including SLIs, SLOs and service reliability
- Strong communication skills and the ability to engage technical and business stakeholders
Desirable Experience- Azure AI Services
- Datadog
- Snowflake
- GitHub Enterprise
- Zero Trust Architecture
- FinOps / Cloud Cost Optimisation
Why Join?This is an opportunity to influence the future direction of a growing Azure estate, working on modern cloud platforms, large-scale automation initiatives and emerging AI technologies. You'll join a collaborative team where engineering quality, innovation and continuous improvement are genuinely valued.
Salary Six Figures + benefits | Hybrid LondonFor a confidential discussion, please get in touch.