Overview
In this role you will design, rollout and support mission-critical network infrastructure across trading, research and GPU compute environments. You will help scale high-bandwidth networks for GPU clusters and deep learning workloads, working closely with systems, research and infrastructure teams. The role emphasizes automation, monitoring and Linux-based troubleshooting to ensure fast, reliable performance. You will contribute to the next phase of network architecture and tooling in a high-performance trading setting. This is a hands-on opportunity to shape a robust, scalable network stack in a demanding environment.
Responsibilities
- Design, rollout and support network infrastructure across trading, research and GPU clusters
- Improve mission-critical, high-bandwidth networking for GPU and compute environments
- Design scalable networks for deep learning and research workloads
- Troubleshoot network issues across Linux-based systems
- Build automation for deployment, monitoring and operations
- Collaborate with systems, research and infrastructure teams
- Support evolution of network architecture and tooling
Key requirements
- Strong experience deploying and supporting production networks
- Good understanding of data centre networking
- Experience with BGP, OSPF and multicast
- Strong Linux or Unix command line skills
- Scripting experience in Python, Bash, Go or similar
- Good troubleshooting skills across network and host-level issues
- Clear communication and strong ownership
- Clear communication
- Strong ownership
- Collaborative mindset
- BGP
- OSPF
- multicast
…
