Network Engineer
Listed on 2026-10-06
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, Network Engineer, IT Infrastructure
Building World Class Teams - MLOps | LLM Inference | HPC | AI - US, APAC and Europe
UK Remote
Are you passionate about Data Centre builds and large-scale GPU infrastructure projects? Do you thrive in a fast-paced, high-growth environment where your work has a direct impact on business outcomes? If so, this could be the role for you!
Our company provides a GPU cloud engineered for AI. We deliver cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. We enable AI-focused companies to achieve superior results by reducing the complexity of AI development. Our infrastructure bolsters technical capabilities and directly supports strategic business outcomes, including cost management, rapid innovation, and environmental responsibility.
Our Engineering team plays a critical role in driving the delivery of our GPU infrastructure. If you're passionate about data centre architecture and thrive in high-performance computing environments, then please apply.
Why Join Us?We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. You’ll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you’ll be contributing to building the technology that powers the future.
AboutThe Role
Network engineers on our team are responsible for the design, deployment, and ongoing operation of all networking services that underpin both the internal management platform and the customer-facing cloud infrastructure. This includes internet transit, WAN connectivity, and DC networking. You will act as a 3/4th line escalation point for the support organisation.
What You’ll Be Doing- Designing, deploying, and operating large-scale HPC clusters and GPU-based compute environments
- Creating and maintaining hardware architectures, including BOMs, rack elevations, and reference designs
- Implementing and maintaining HPC scheduling and workload management systems (e.g., Slurm)
- Designing and optimising Infini Band and Ethernet network topologies (Fat Tree, Dragonfly, rail-optimized configurations)
- Working with deployment teams to ensure cluster builds align with architectural specifications
- Automating provisioning, configuration, and operations of multi-vendor HPC hardware and software stacks
- Troubleshooting and tuning cluster performance across compute, storage, and interconnect layers
- Collaborating with software, infrastructure, and data center teams to ensure seamless integration of HPC environments
- Proven experience designing, deploying, and operating HPC or large-scale compute clusters
- Strong knowledge of Slurm or similar workload management systems (e.g., PBS, LSF)
- Proven experience in Infini Band networking design and operations, including subnet management, QoS, RDMA, and performance tuning
- Experience with high-speed Ethernet networks and associated protocols (e.g., VLAN, LACP, BGP, OSPF, EVPN, VXLAN)
- Familiarity with HPC network topologies such as Fat Tree or Dragonfly
- Experience creating hardware BOMs, rack layouts, and reference architectures for compute deployments
- Strong scripting skills in Python and/or Bash for automation and orchestration
- Solid understanding of optics, cabling, and physical layer design considerations for HPC and GPU cluster environments
- Strong analytical, troubleshooting, and documentation skills
- A collaborative mindset and passion for building high-performance, scalable infrastructure
- Proactive and self-motivated, with a strong sense of ownership
- Thrives in a fast-paced, dynamic, and high-growth environment
- Collaborative team player with a passion for delivering outstanding…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).