openstack HPC stackhpc community news AI azimuth ironic monitoring deployment baremetal kolla-ansible ansible networking kayobe kolla infrastructure kubernetes slurm monasca scheduling newsletter mpi virtualisation beegfs cloudkitty sriov cluster training ciuk operations recruitment network overcloud slinky ceph data github ucx galaxy cloud-init storage gpu grafana cuda vm kata ci roce runc dnf
StackHPC develops open source cloud infrastructure for research computing, HPC and AI. We are proud to work with clients tackling the demanding challenges of modern scientific computing. StackHPC contributes actively to open source projects and communities. Our team keeps a close eye on the latest developments and technologies in order to provide up to date and competitive solutions for a range of client requirements. We foster a collaborative and inclusive work environment where every team member can thrive and grow.
Job Description
We are seeking a motivated and organised Cloud Engineer to join our dynamic team. The ideal candidate will have experience of Kubernetes and a keen interest in continuous integration, continuous delivery (CI/CD), and infrastructure automation. Cloud Engineers play a key role in maintaining and optimising our development and deployment pipelines, ensuring the reliability and scalability of our systems.
While this role is ideally based on-site in Bristol, we are open to fully remote working arrangements for candidates based in the UK or France who demonstrate exceptional skills and significant industry experience.
Key Responsibilities
- Assist in the development, maintenance, and improvement of CI/CD pipelines.
- Work within the development team to integrate new features and enhancements.
- Implement and manage container orchestration using Docker and Kubernetes.
- Troubleshoot and resolve infrastructure and deployment issues.
- Participate in code reviews and provide constructive feedback.
Qualifications
Core Requirements
- Bachelor’s degree in Computer Science, Information Technology, or a related field.
- Familiarity with Linux.
- Basic understanding of DevOps principles and practices.
- Some experience with Kubernetes and containerisation.
- Familiarity with CI/CD tools (e.g., GitHub workflows, GitLab pipelines).
- Knowledge of scripting languages (e.g., Bash, Python).
- Strong problem-solving skills and attention to detail.
- Excellent communication and teamwork abilities.
- Eagerness to learn and adapt to new technologies and challenges.
- Experience with cloud platforms (e.g., OpenStack, AWS, Azure, GCP).
Additional Assets
- Understanding of infrastructure-as-code tools (e.g., Terraform, Ansible).
- Exposure to monitoring and logging tools (e.g., Prometheus, Grafana, OpenSearch).
- Experience using or operating an OpenStack cloud.
- Knowledge of high performance computing platforms and concepts.
We work with and deploy a wide range of OpenStack technologies and services. Experience with the following additional technologies is always beneficial:
- O/S: Rocky Linux, Ubuntu
- Storage: Ceph, NFS
- Networking: Ethernet, Infiniband, SRIOV
- Compute and Workloads: Kubernetes, Slurm
Our technology stack is constantly evolving to meet market and user requirements. So we will help you get up to speed as required.
- Flexible working hours and discretionary remote working / flexible working practices for senior-level expertise.
- Pension contribution.
- Support for travel to conferences and delivering presentations.
- Learning and employee development is a priority, including work time dedicated to R&D and technology.
We have an existing recruiter relationship, so there is no need for recruiters to contact us about this role.
#J-18808-Ljbffr…
