logo of nvidia

英伟达AI Computing Development Engineer, TensorRT and TensorRT-LLM AIGV

社招全职状态:招聘

任职要求


• Masters or higher degree in Computer Engineering, Computer Science, Applied Mathematics, or related computing-focused field (or equivalent experience)
• Strong Python or C/C++ programming and software design experience, including debugging, performance profiling, and test design
• 2+ years working experience
• Strong curiosity about artificial intelligence and familiarity with the latest developments in deep learning — including gener…
登录查看完整任职要求
微信扫码,1秒登录

工作职责


• Design and develop robust inferencing software (TensorRT/TensorRT-LLM) optimized for functionality and performance across platforms
• Perform performance analysis, optimization, and tuning of deep learning inference workloads
• Track and integrate academic and industry advancements in AI and feature-update TensorRT/TensorRT-LLM accordingly
• Provide feedback into architecture and hardware design and development
• Collaborate across hardware, software, and research teams to shape the direction of machine learning inferencing across NVIDIA platforms
• Own and deliver technical work with scope based on experience, ranging from complex features to substantial parts of larger projects, with increasing independence and technical leadership over time
• Publish key technical results at leading scientific and engineering conferences
包括英文材料
相关职位

logo of nvidia
社招

• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and feature update TensorRT • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences

更新于 2025-11-03上海
logo of nvidia
社招

• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and feature update TensorRT • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences

更新于 2026-02-04上海
logo of nvidia
社招

• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and feature update TensorRT • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences

更新于 2026-06-11上海
logo of nvidia
社招

• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and large language models • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences

更新于 2026-06-16上海|北京