Overview
As Lead Platform Engineer, you own and evolve the cloud platform to support a growing data business. You will secure, scale and cost-optimize the Azure-based estate while driving automation and IaC to modernize infrastructure. You’ll lead a cloud separation and migration program spanning Azure and GCP, shaping platform standards and developer experience. This is a hands-on, influential role with cross-functional collaboration and opportunities to define architecture and tooling. You’ll work in a collaborative, hybrid setup with visible impact on reliability and data platform capabilities.
Pay / Benefits
- Private healthcare after probation
- Charity donation scheme
- Work-from-anywhere week
- Hybrid working
- Regular team events
- Flexible and impactful role
Responsibilities
- Own and manage the Azure cloud estate with a focus on security, reliability and cost efficiency
- Lead adoption of Terraform and Infrastructure as Code, bringing existing infrastructure under management
- Own and operate Kubernetes workloads including deployments, ingress, scaling, upgrades, troubleshooting and day-to-day platform operations
- Build and maintain CI/CD pipelines using Azure DevOps
- Support cloud migration and separation projects across Azure and GCP
- Design, automate and test disaster recovery and business continuity solutions
- Improve monitoring, alerting and observability across the platform
- Manage cloud security, identity, secrets and certificates
- Provide technical guidance on Azure architecture and platform engineering
- Troubleshoot complex infrastructure issues and act as senior technical escalation point
- Collaborate with development and technology teams to improve platform reliability and developer experience
Key requirements
- Azure cloud engineering experience with focus on security, reliability and cost efficiency
- Proven experience with Terraform and Infrastructure as Code
- Hands-on Kubernetes operations and management
- CI/CD with Azure DevOps
- Experience with cloud migrations, especially Azure and GCP
- Disaster recovery and business continuity design and testing
- Monitoring and observability improvements
- Cloud security, identity, secrets and certificates management
- Strong troubleshooting and senior escalation capabilities
- Collaboration with cross-functional teams to improve platform reliability
- collaboration
- leadership
- problem-solving
- Azure
- GCP
- Terraform
…
