英伟达AI Computing Development Engineer, TensorRT and TensorRT-LLM AIGV
任职要求
• Masters or higher degree in Computer Engineering, Computer Science, Applied Mathematics, or related computing-focused field (or equivalent experience) • Strong Python or C/C++ programming and software design experience, including debugging, performance profiling, and test design • 2+ years working experience • Strong curiosity about artificial intelligence and familiarity with the latest developments in deep learning — including gener…
工作职责
• Design and develop robust inferencing software (TensorRT/TensorRT-LLM) optimized for functionality and performance across platforms • Perform performance analysis, optimization, and tuning of deep learning inference workloads • Track and integrate academic and industry advancements in AI and feature-update TensorRT/TensorRT-LLM accordingly • Provide feedback into architecture and hardware design and development • Collaborate across hardware, software, and research teams to shape the direction of machine learning inferencing across NVIDIA platforms • Own and deliver technical work with scope based on experience, ranging from complex features to substantial parts of larger projects, with increasing independence and technical leadership over time • Publish key technical results at leading scientific and engineering conferences
• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and feature update TensorRT • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences
• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and feature update TensorRT • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences
• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and feature update TensorRT • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences
• Craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance • Performance analysis, optimization and tuning • Closely follow academic developments in the field of artificial intelligence and large language models • Provide feedback into the architecture and hardware design and development • Collaborate across the company to guide the direction of machine learning inferencing, working with software, research and product teams • Publish key results in scientific conferences