Senior Research Engineer - Enterprise Products
Listed on 2026-07-15
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer
Overview
We are now looking for a Senior Research Engineer passionate about Generative AI inference. Are you excited to change the way people infuse AI into products and services? NVIDIA is at the forefront of generative AI models, from language to images. NVIDIA provides building blocks to democratize AI and make generative AI easy to develop, integrate, and deploy. Our team is dedicated to developing optimized inferencing technologies to support our growing generative AI needs.
We collaborate with research teams, engineers, and the open-source community.
- Design and evaluate routing policies for LLM traffic to best use a mixture of model systems.
- Build and run agentic benchmarks (e.g., Terminal-Bench) to measure algorithm quality, and turn results into calibration data and routing profiles.
- Ship to an open-source repo: design docs, code review, docs, and community contributions.
- Collaborate with engineering teams across NVIDIA to ensure our software integrates seamlessly up and down the NVIDIA accelerated serving stack.
- Bachelor's or Master’s degree in Computer Science or equivalent experience.
- 8+ years of industry experience in Deep Learning frameworks (PyTorch or Tensor Flow).
- Experience designing or running LLM evaluations/benchmarks — ideally agentic ones — and drawing statistically sound conclusions from them.
- Understanding of modern techniques in Machine Learning, Deep Neural Networks, Natural Language Processing, or Speech Recognition.
- Empirical research mindset: forming hypotheses about new algorithms, running calibrations, iterating on results.
- Strong communication and interpersonal skills, with the ability to work in a dynamic and distributed team. A history of mentoring junior engineers and interns is a plus.
- A desire to constantly grow and learn new things.
- Strong computer science fundamentals — algorithms and data structures, computational complexity, parallel and distributed computing, system software.
- Experience architecting or developing large-scale distributed systems for deep learning.
- Agentic benchmark creation and publications.
- Knowledge of CPU and/or GPU architecture.
- GPU programming (CUDA).
- Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 192,000 USD - 304,750 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.
- You will also be eligible for equity and benefits.
Applications for this job will be accepted at least until July 14, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.
EEO StatementNVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. We highly value diversity in our current and future employees, and we do not discriminate on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
#J-18808-Ljbffr(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).