英伟达AI Computing Software Development Intern - 2026
任职要求
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. What You’ll be doing: As an intern, you’ll focus on one of three specialized tracks:• TensorRT‑LLM – Inference Optimization (Python / PyTorch): Build and enhance high‑performance LLM inference pipelines. Analyze and optimize model execution, scalability, and memory use. Collaborate across framework and research teams to deliver efficient multi‑GPU model serving. • TensorRT Compiler – Graph Optimization (C++): Work on the TensorRT compiler backend to improve graph transformations and code generation for NVIDIA GPUs. Develop compiler optimization passes, refine operato…
工作职责
N/A
• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and large language models • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences
• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and feature update TensorRT • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences
• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and feature update TensorRT • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences
• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and feature update TensorRT • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences