Senior Software Developer, AI Networking
Listed on 2026-07-19
-
Software Development
AI Engineer (Applied/Software), Software Engineer, DevOps
NVIDIA is changing the world of AI Networking with groundbreaking technology. We are excited to be adding an AI Networking Software Developer to our AI Networking SW development and codesign team. We are working with the latest NVIDIA hardware and technologies. We do full stack benchmarking for Data Center scale systems for AI training/inference and lower level benchmarks. We strive for automation and develop many tools in-house yet adopt community accepted practices and frameworks.
Moreover we give back to community developing our own tools in public Git Hub repositories. Our goal is to ensure that large-scale systems deliver expected performance in practice, not just on paper, by uncovering bottlenecks and driving continuous improvements.
- Developing AI networking communication frameworks and applications running in production on the world’s largest supercomputers and data centers.
- Develop production tools and benchmarks used by multiple teams inside and outside NVIDIA.
- Enable new AI models within our benchmarking infrastructure and deliver insights through end-to-end analysis of large-scale workloads across hardware and software stacks.
- Design and implement automation systems, including large-scale parameter search to identify optimal configurations across complex systems.
- Collaborate closely with networking and hardware teams to co-design new features and software interfaces in a fast-paced, evolving environment.
- B.Sc., M.Sc degree in Computer Science / Software engineering or equivalent experience.
- 5+ years of experience.
- Professional Python development experience. We seek individuals who build maintainable, long-lived tools that do not impose a heavy burden on the team in terms of maintenance.
- Solid Linux expertise and passion for working extensively in command-line environments.
- Ability to work across a broad and evolving stack, with a strong drive to learn—from hardware and networking up to large-scale AI systems running across entire clusters.
- Knowledge and/or experience with modern AI ecosystem:
PyTorch, LLMs, inference and training. - Familiarity with cluster orchestration systems such as Slurm or Kubernetes.
- Knowledge in MPI and HPC, Infini Band, Ethernet and Networking.
- Experience in performance optimizations.
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is $152,000 USD - $241,500 USD for Level 3 and $184,000 USD - $287,500 USD for Level 4.
You will also be eligible for equity and benefits.
Equal Opportunity EmployerNVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
#J-18808-Ljbffr(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).