Senior System Software Engineer – Dynamo Tools
Listed on 2026-07-23
-
Software Development
AI Engineer (Applied/Software), AI Reliability/ Performance Engineer, Cloud Engineer - Software, Software Engineer
Overview
We are seeking a Senior System Software Engineer to own and advance the AI-Perf analysis, NVIDIA’s flagship framework for benchmarking, experimentation, and analysis of LLMs, Generative AI, and deep learning inference workloads. In this role, you’ll combine systems research, distributed systems engineering, and applied AI, enabling reproducible performance evaluation, influencing internal platforms, and providing tooling that empowers researchers and engineers globally.
ResponsibilitiesLead the design, development, and roadmap of AI-Perf, defining benchmarking methodologies, performance metrics, and reproducible experimental workflows.
Build scalable and high-performance features to measure latency, throughput, and efficiency across AI models and distributed systems.
Partner with AI researchers, platform teams, and engineers to translate experimental challenges into robust, user-friendly performance tooling.
Integrate AI-Perf with the Dynamo Inference Stack, other NVIDIA inference stacks, and open-source inference frameworks, delivering end-to-end performance insights for researchers and production users.
Qualifications- Bachelor’s, Master’s, or PhD in Computer Science, Computer Engineering, or related field—or equivalent experience.
- 3+ years of experience in systems software, distributed performance engineering, or AI infrastructure research.
- Expert-level Python skills, including profiling, optimization, automation, and debugging of complex systems.
- Deep knowledge of distributed systems concepts, including scalability, concurrency, fault tolerance, and performance trade‑offs.
- Experience designing or maintaining performance benchmarking frameworks or tooling for AI/ML systems.
- Hands‑on experience with LLMs and deep learning frameworks such as PyTorch, Tensor Flow, TensorRT, or ONNX Runtime.
- Contributions to open-source or research projects in AI performance, infrastructure, or distributed systems.
- Experience running large‑scale inference experiments across cloud and on‑prem environments (AWS, Azure, GCP, bare metal).
Base salary range: $152,000 – $241,500 for Level 3 and $184,000 – $287,500 for Level
4. Eligible for equity and additional benefits. Competitive compensation package.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
NVIDIA uses AI tools in its recruiting processes.
#J-18808-Ljbffr(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).