Job Description & How to Apply Below
Join NVIDIA as a Senior Engineer focused on AI Infrastructure. Leverage your extensive background in GPU technologies, high-performance computing, and systems programming to push boundaries.
In this dynamic role, you will apply over a decade of software engineering experience to enhance NVIDIA's AI capabilities. Your responsibilities will include building an automated platform for data center management, improving system monitoring, and ensuring high availability for GPU assets.
You'll work with a talented team to innovate and solve complex problems in machine learning operations.
Key Responsibilities:
• Build automated solutions for GPU asset management
• Ensure industry-leading reliability and scalability of systems
• Implement health management solutions using telemetry
• Develop software for NVLINK management across clusters
• Collaborate across teams for seamless integration
Requirements:
• Over 10 years of experience with production systems
• Bachelor’s degree in Computer Science or equivalent
• Expertise in systems programming languages like Go, Python
• Advanced Linux system management capabilities
• Familiar with Kubernetes and SLURM for cluster management
Make an impact at NVIDIA by enhancing AI infrastructure and deploying innovative solutions.
#J-18808-Ljbffr
Position Requirements
10+ Years
work experience
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
Search for further Jobs Here:
×