Deep Learning Architect, LLM Inference - New College Grad 2026 at NVIDIA

**Who this is for** This position is tailored for a 2026 college graduate with a deep interest in performance benchmarking, workload characterization, and the a

Work type: onsite

Location: US, CA, Santa Clara

Salary: $124,000 – $241,500/yr

Type: Full-time

Summary

**Who this is for** This position is tailored for a 2026 college graduate with a deep interest in performance benchmarking, workload characterization, and the architectural optimization of Large Language Models. **Key highlights** You will play a critical role in establishing benchmarking methodologies and performance tools to ensure NVIDIA’s hardware and software ecosystem maintains industry leadership in inference speed and efficiency. **You might be a good fit if you...** - Hold a Master’s or PhD in Computer Science or related fields. - Have practical experience with LLM inference servers like vLLM or TRT-LLM. - Possess a solid understanding of CPU/GPU microarchitecture and performance bottlenecks. - Are proficient with AI coding agents and automated profiling tools.

Job Description

We are now looking for a Deep Learning Architect, LLM Inference!

NVIDIA is at the forefront of the generative AI revolution. The Inference Benchmarking (IB) team specifically focuses on inference server performance optimization for Large Language Models (LLMs). If you're passionate about pushing the boundaries of GPU hardware and software performance and understand terms like disaggregated serving, data parallel attention, MoE, Qwen3.5, DeepSeek, GPT-OSS, then this is a great role for you!

What you'll be doing:










What we need to see:









Ways to stand out from the crowd:




NVIDIA is widely considered to be one of the technology world's most desirable employers. We have a team of highly skilled and motivated individuals who excel in their work. If you have a proactive and independent approach, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 124,000 USD - 195,500 USD for Level 2, and 152,000 USD - 241,500 USD for Level 3.

You will also be eligible for equity and [benefits](https://www.nvidia.com/en-us/benefits/).

Applications for this job will be accepted at least until April 26, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

View this job on nocollar jobs