×
Register Here to Apply for Jobs or Post Jobs. X

GPU Network Engineer

Job in Santa Clara, Santa Clara County, California, 95053, USA
Listing for: Blue Signal Search
Full Time position
Listed on 2026-08-22
Job specializations:
  • IT/Tech
    Network Engineer, Systems Engineer, Cloud Computing: Infrastructure & Operations
Salary/Wage Range or Industry Benchmark: 180000 - 240000 USD Yearly USD 180000.00 240000.00 YEAR
Job Description & How to Apply Below

An innovative technology organization at the forefront of AI infrastructure is seeking a GPU Network Engineer to help design and operate the high-performance networking backbone powering advanced GPU computing environments. This opportunity is ideal for an engineer who enjoys solving complex networking challenges, optimizing ultra low latency communication, and building highly scalable infrastructure supporting large AI training and inference workloads. You will work alongside experienced infrastructure and compute professionals while influencing the architecture of next generation GPU clusters.

What

You Will Do
  • Design, deploy, and optimize high performance Infini Band (NDR and XDR) and RoCEv2 fabrics supporting large scale GPU and AI workloads.
  • Build scalable CLOS and ECMP network architectures that deliver low latency, high bandwidth east west traffic across distributed GPU environments.
  • Implement and support EVPN, VXLAN, Layer 2, and Layer 3 networking for modern data center infrastructure.
  • Configure and maintain Arista, Juniper, and NVIDIA networking platforms while ensuring highly available, resilient network operations.
  • Automate network provisioning, configuration, and lifecycle management using Netconf, Ansible, Terraform, and Net Box.
  • Monitor, troubleshoot, and optimize network performance through packet analysis, telemetry, and root cause analysis.
  • Partner with infrastructure, compute, and platform engineering teams to support AI training, inference, and high performance computing initiatives.
  • Create and maintain technical documentation, operational procedures, and implementation standards while participating in change management and infrastructure improvements.
Required Qualifications
  • Minimum of 5 years of experience supporting enterprise data center or HPC networking environments with direct GPU cluster networking experience.
  • Extensive hands on experience with Infini Band (NDR or XDR) and RoCEv2 networking technologies.
  • Strong expertise designing and supporting east west networking for GPU or AI infrastructure.
  • Experience with CLOS architectures, ECMP routing, EVPN, and VXLAN.
  • Hands on administration of Arista EOS, Juniper Junos, and NVIDIA or Mellanox networking platforms.
  • Experience using Net Box, Netconf, Ansible, and Terraform to automate network deployment and management.
  • Strong troubleshooting skills, packet analysis experience, and network performance optimization capabilities.
  • Excellent communication and technical documentation skills.
Preferred Qualifications
  • Experience optimizing fabrics supporting large scale AI training workloads.
  • Knowledge of adaptive routing, congestion management, RDMA optimization, and lossless networking.
  • Python scripting for infrastructure automation.
  • Industry certifications such as CCNP, CCIE, JNCIP, JNCIE, or equivalent practical experience.
#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary