Cloud Operations Engineer

Company: Kerridge Commercial Systems
Apply for the Cloud Operations Engineer
Location: Tankersley
Job Description:

Overview

As a Cloud Operations Engineer, you will own the day-to-day operation and improvement of cloud-based products, ensuring security, reliability, and performance. You’ll monitor incidents, manage IaC-driven infrastructure, and evolve CI/CD pipelines to enable automated deployments. You’ll work with cross-functional teams to meet SLOs, support disaster recovery, and uphold security and compliance standards. This role offers a chance to advance ops maturity, embrace AI-enabled tooling, and contribute to a fast-growing, global product ecosystem.

Pay / Benefits

  • hybrid work policy (3 days in office, 2 remote)
  • on-call compensation
  • global team collaboration
  • opportunities to work with AI-enabled technologies

Responsibilities

  • Own day-to-day operation, availability, resilience, and continuous improvement of cloud-based products and services
  • Monitor, diagnose, and resolve incidents across compute, storage, network, and identity platforms
  • Manage cloud infrastructure using Infrastructure as Code (IaC) tools
  • Design, build, and support CI/CD pipelines for automated deployments
  • Ensure services meet defined SLOs and utilize observability data for decisions
  • Perform backup, DR, and failover activities with tested RTO/RPO
  • Implement security controls, IAM, patching, and vulnerability remediation
  • Manage secrets and privileged access in line with governance
  • Ensure compliance with ISO 27001, SOC 2, GDPR and related standards
  • Identify recurring issues and implement preventative measures
  • Develop automation to reduce manual effort and improve reliability
  • Apply FinOps to optimise cloud costs
  • Maintain runbooks and documentation
  • Participate in on-call rotations and respond to incidents
  • Collaborate with Ops, Governance, Security, and Eng Engineering teams
  • Mentor junior engineers and promote best practices
  • Support capacity planning, performance optimization, and cost management

Key requirements

  • Solid hands-on experience operating a cloud platform in production with depth in either Microsoft Azure or OCI
  • Incident resolution and troubleshooting across compute, storage, network, and identity
  • Experience with Infrastructure as Code (e.g., Terraform or Bicep) and CI/CD pipelines
  • Scripting proficiency (PowerShell, Bash) for automation
  • IAM, patching, vulnerability remediation and safe secrets management
  • Experience working to SLAs/SLOs and in compliance controls (ISO 27001, SOC 2, GDPR)
  • Willingness to participate in a compensated on-call rota
  • collaboration
  • curiosity about AI and learning mindset
  • ability to communicate across teams
  • Azure or OCI cloud platform expertise
  • Terraform or Bicep
  • PowerShell or Bash scripting

…

Posted: October 1st, 2026