DevOps / SRE Engineer

Company: Bumper
Apply for the DevOps / SRE Engineer
Location: London
Job Description:

  • Design, build and maintain secure, scalable cloud infrastructure across AWS and Azure
  • Manage and enhance our Kubernetes (EKS) platform to support reliable, modern applications
  • Develop and maintain Infrastructure as Code using Terraform and Helm
  • Improve and support CI/CD pipelines using Argo Workflows, ArgoCD and GitHub Actions
  • Lead and participate in incident response, including on-call activities and major incident coordination
  • Drive high-quality monitoring, alerting and observability across metrics, logs and traces
  • Conduct and support blameless post-incident reviews, ensuring follow-up actions are delivered
  • Define and implement SLIs/SLOs to improve service reliability and operational excellence
  • Collaborate with engineering teams to embed best practices and improve developer experience
  • Contribute to automation, tooling, and continuous improvements that reduce toil and increase platform resilience

Benefits

  • Share Options: Our people drive our success. We reward that hard work with ownership in the company.
  • Bumper Flex: We all work in different ways. Bumper Flex brings you the best of both worlds, home & office.
  • Annual Retreat: Remote working is fab. But nothing beats getting to know each other in person.
  • Bumper Allowance: £250 a year to spend on your personal wellbeing and development, simple as that!
  • Relax and Recharge: 26 days of paid leave each year to do whatever it is that recharges your batteries.
  • Family Leave: 4 months’ paid leave for primary carers & 1 month for secondary carers. Your family comes first.
  • Colleague Discount: Discount on Bumper loans for you, your friends, and family – Boom!
  • Colleague Assistance Programme: 24/7 support for your wellbeing
  • Bumper Foundation: We’re carbon neutral & offer paid leave each year to give back, in whatever way you choose.

Solid experience with Terraform and IaC automationExcellent communication and calmness under pressureDemonstrated use of GenAI tools (ChatGPT, GitHub Copilot, Claude) in engineering workflowsAbility to diagnose and fix complex distributed systems issuesStrong hands-on experience running Kubernetes in productionExperience participating in or managing production incidents and on-callExperience with AWS and/or Azure cloud platformsA passion for automation and reducing toilProven experience in DevOps, SRE, or Platform Engineering rolesStrong grasp of monitoring, alerting, and observability principlesFamiliarity with Argo Workflows / ArgoCDExperience defining SLIs/SLOs at scaleBackground in Platform EngineeringExperience with Grafana Cloud, IRM or other incident management toolsExperience improving or redesigning incident management processes

#J-18808-Ljbffr…

Posted: September 28th, 2026