Windows AI Engineering Intern - Fall 2025

US, CA, Santa Clara, United States

NVIDIA

NVIDIA on grafiikkasuorittimen keksijä, jonka kehittämät edistysaskeleet vievät eteenpäin tekoälyn, suurteholaskennan.

View all jobs at NVIDIA

Apply now Apply later

At NVIDIA, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world!

As a Windows AI Engineering Intern, you will be instrumental in developing inference runtimes, optimizing GenAI pipelines and inference backends, and devising algorithms that flawlessly incorporate AI into games and applications for Windows. This position requires a unique blend of AI knowledge and system software skills, making it an excellent fit for individuals who have a strong desire to shape the future of computing!

What You’ll Be Doing:

  • Partnering with NVIDIA software, research, architecture, and product teams, aligning strategies and technical needs for fostering the ecosystem of AI on a Windows RTX PC.

  • Performing in-depth analysis and optimization of AI models, AI frameworks, data processing pipelines, and inference backends to ensure the best performance on current and next-generation GPU architectures.

  • Identifying and implementing compute and memory optimizations across the full AI inference stack on RTX Windows PC.

  • Developing model compression and fine-tuning techniques to reduce resource consumption and improve performance, enabling efficient deployment and better user experience.

  • Designing and implementing an optimized framework for running AI NPCs in gaming applications as part of the NVIDIA ACE Platform.

  • Collaborating with Microsoft to drive advancements in APIs, AI frameworks, and platforms for developing and deploying AI inferencing applications.

  • Ensuring the effective deployment of directed tests through collaboration with the automation team, thereby ensuring the robustness of automated testing.

What We Need to See:

  • Pursuing BS, MS, or PhD in Computer Science, Software Engineering, Mathematics, or a related field.

  • Experience with AI inferencing pipelines and applications using ML/DL frameworks like PyTorch, ONNX Runtime preferred.

  • Excellent C++ programming and debugging skills with a strong understanding of data structures and algorithms.

  • Strong analytical and problem-solving abilities, with the capacity to multitask effectively in a dynamic environment.

Ways To Stand Out From The Crowd:

  • Understanding of modern techniques in Machine Learning, Deep Neural Networks, and Generative AI with relevant contributions to major open-source projects will be a plus.

  • Proficiency in lower-level system/GPU programming, CUDA, and developing high-performance systems.

  • Hands-on experience with building applications using graphics APIs like OpenGL, DirectX, Vulkan, etc.

  • Consistent track record of delivering end-to-end products with geographically distributed teams in multinational product companies.

Are you dedicated, upbeat and dynamic with excellent analytical ability? Are you an engineer passionate and highly motivated about solving complex problems? If so, you may be a perfect fit for NVIDIA!

The hourly rate for our interns is 18 USD - 71 USD. Our internship hourly rates are a standard pay determined based on the position and your location, year in school, degree, and experience.

You will also be eligible for Intern benefits. NVIDIA accepts applications on an ongoing basis. ​

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Apply now Apply later
Job stats:  4  1  0

Tags: APIs Architecture Computer Science CUDA Engineering Generative AI GPU Machine Learning Mathematics ONNX Open Source PhD Pipelines PyTorch Research Testing Vulkan

Perks/benefits: Career development

Region: North America
Country: United States

More jobs like this