AI Computing Software Intern, GPU Kernel Libraries - 2027

nvidiaΒ· CN05 NVIDIA Shanghai WFOE
Apply Now β†—
πŸ“ China, BeijingπŸ“ China, ShanghaiFull time

About this role

We are now looking for Compute/DL Architecture Performance Optimization Interns in our group! Are you passionate about exploring computer architectures for deep learning? Do you enjoy working at the intersection of hardware and software? NVIDIA is looking for world-class programmers and performance architects who like to continuously explore and mine the ultimate performance of each operator (or fusion operator) in deep learning networks, design and develop scalable modular infrastructure that can ship these highly optimized operators to different NVIDIA software libraries for training and inference.

What you'll be doing:

  • Develop high performance operators on NVIDIA GPUs for cuBLAS, TensorRT, cuDNN, cuSparse and cuTensor libraries

  • Analyze the performance of various GPU kernels on existing/new architecture, identify bottlenecks and propose creative solutions to improve them

  • Design and develop software for kernel authoring and shipping

  • Adopt cutting-edge AI technologies in GPU kernel or similar development workflow

What we need to see:

  • Pursuing a B.S., M.S., or PhD degree in computer science (or similar)

  • Strong programming skills in C/C++ and Python development

  • Familiar with GPU programming model and CUDA

  • Good understanding about compiler technologies and experience with LLVM and MLIR

  • Excellent problem solving skills, good communication and teamwork

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!

Frequently Asked Questions

Is the salary disclosed for the AI Computing Software Intern, GPU Kernel Libraries - 2027 position at nvidia?
The salary for this AI Computing Software Intern, GPU Kernel Libraries - 2027 role at nvidia is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Where is the AI Computing Software Intern, GPU Kernel Libraries - 2027 position at nvidia located?
This AI Computing Software Intern, GPU Kernel Libraries - 2027 role at nvidia is based in China, Beijing, China, Shanghai. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
Is the AI Computing Software Intern, GPU Kernel Libraries - 2027 role at nvidia full-time or part-time?
This is listed as a Full time position. It is posted as a AI Computing Software Intern, GPU Kernel Libraries - 2027 role in the CN05 NVIDIA Shanghai WFOE department at nvidia.
Which team or department does the AI Computing Software Intern, GPU Kernel Libraries - 2027 at nvidia belong to?
This AI Computing Software Intern, GPU Kernel Libraries - 2027 position is part of the CN05 NVIDIA Shanghai WFOE department at nvidia. See the full job description for more information about the team structure and responsibilities.
How do I apply for the AI Computing Software Intern, GPU Kernel Libraries - 2027 position at nvidia?
Click the "Apply Now" button on this page. You will be redirected to nvidia's official application portal hosted on workday where you can submit your application directly.
When was the AI Computing Software Intern, GPU Kernel Libraries - 2027 job at nvidia posted?
This AI Computing Software Intern, GPU Kernel Libraries - 2027 position at nvidia was posted on Sep 23, 2026. Apply as soon as possible β€” early applications are often reviewed first.
AI Computing Software Intern, GPU Kernel Libraries - 2027
nvidia
Apply for this role β†—

You'll be redirected to nvidia's official application page on Workday.