Senior Staff AI Infrastructure Engineer

Company: Boehringer Ingelheim
Apply for the Senior Staff AI Infrastructure Engineer
Location: London
Job Description:

Overview

In this senior role, you own and evolve the AI Accelerator’s infrastructure strategy, ensuring compute, storage and platform support for training, fine-tuning and serving biomedical models at scale. You lead multi-layer architecture, set technical direction, and partner with IT and cloud providers to maximize efficiency and portability. You’ll mentor engineers and align infra with research priorities, enabling secure, compliant AI workloads. Join a mission-driven team building explainable AI to advance disease understanding.

Responsibilities

  • Define and evolve a holistic reference architecture spanning storage, compute, environments and platform stack
  • Set technical direction and roadmap for AI infrastructure aligned to CI roadmaps and research priorities
  • Collaborate with Data Excellence to advance Trusted Research Environment (TRE) for secure AI workloads
  • Shape evolution of enterprise infrastructure with IT to support large-scale AI and optimize cost
  • Establish onboarding and mentoring practices; act as senior escalation point for complex infra issues

Key requirements

  • PhD or MSc in a STEM subject
  • Extensive experience as a senior staff-level infrastructure, platform or systems engineer for compute-intensive workloads
  • Deep expertise in AI compute infrastructure: GPU compute, storage, networking, Kubernetes, IaC (Terraform)
  • Hybrid on-prem/cloud architectures with cost optimization and portability
  • Experience enabling large-scale AI workloads with distributed training
  • Security, access control, resilience and ITIL-like service management
  • Strong collaboration and influencing skills; ability to explain complex concepts to diverse stakeholders
  • Mentoring or technical leadership of engineers and guiding engineering principles
  • collaboration and influencing across technical and non-technical stakeholders
  • mentoring and leadership
  • clear communication of complex concepts
  • GPU compute
  • high-performance compute
  • storage and file systems

…

Posted: September 30th, 2026