×
Register Here to Apply for Jobs or Post Jobs. X

Senior HPC Storage Engineer

Job in Santa Clara, Santa Clara County, California, 95053, USA
Listing for: NVIDIA Gruppe
Full Time position
Listed on 2026-09-12
Job specializations:
  • IT/Tech
    Systems Engineer, Cloud Computing: Infrastructure & Operations
Salary/Wage Range or Industry Benchmark: 184000 - 287500 USD Yearly USD 184000.00 287500.00 YEAR
Job Description & How to Apply Below

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world.

Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIA, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

As a member of the HW Infrastructure Storage Strategy team, you will provide leadership in the research, design and implementation of groundbreaking fast storage solutions to enable runs of demanding high‑performance computing and computationally intensive workloads. You are expected to identify architectural changes encompassing file, block, and object storage, to cater to the scaling and performance requirements of an expanding cloud infrastructure.

You will help shape next‑generation storage solutions, address strategic challenges in storage design for large‑scale high‑performance workloads, evolve our private/public cloud strategy, perform capacity modelling, and plan growth across our global computing environment.

What you'll be doing
  • Research and analyze existing internal distributed storage services.
  • Research, design, and implement scalable, next‑gen distributed storage services for HPC workloads, optimizing both performance and cost‑effectiveness to meet NVIDIA’s growing infrastructure needs.
  • Develop tooling to automate management of large‑scale infrastructure environments, automate operational monitoring and alerting, and enable self‑service consumption of resources.
  • Detail general procedures and practices, perform technology evaluations related to distributed file systems.
  • Collaborate across teams to understand developers' workflows and capture their infrastructure requirements.
  • Influence and guide methodologies for building, testing, and deploying applications to ensure efficient performance and resource utilization.
  • Support researchers to run their flows on our clusters, including performance analysis and optimizations of deep‑learning workflows.
  • Perform root cause analysis and suggest corrective action for problems at large and small scales.
What we need to see
  • Bachelor’s degree in Computer Science, Electrical Engineering or related field or equivalent experience.
  • 8+ years of experience designing and/or operating large‑scale storage infrastructure.
  • Experience analyzing and tuning storage performance for a variety of workloads.
  • Proficient in CentOS/RHEL and/or Ubuntu Linux distros including Python programming and Bash scripting.
  • In‑depth understanding of container technologies like Docker and Enroot.
Ways to stand out from the crowd
  • Distributed Storage Expertise – Extensive experience with parallel and distributed file systems (Ceph, Weka.io, Vast, Lustre, GPFS) and Linux storage kernel development.
  • GPU & AI Infrastructure – Proficient with NVIDIA GPUs, CUDA programming, and NCCL, including performance benchmarking via MLPerf.
  • Hardware & Storage Engineering – Deep familiarity with storage hardware (HDDs, SSDs, NVMe), enclosures, and specialized appliances like Network Appliance.
  • Advanced Networking – Strong background in Software‑Defined Networking (SDN) and high‑performance networking for AI/HPC clusters.
  • Deep Learning Frameworks – Practical experience applying industry‑standard frameworks, specifically PyTorch and Tensor Flow.
Benefits and Compensation

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is  184,000 USD –  287,500 USD for Level 4, and  224,000 USD –  356,500 USD for Level 5. You will also be eligible for equity and benefits.

Equal Opportunity Employer

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. We do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary