高德地图高德-AI Infra/模型推理工程师/专家-推理加速方向-视觉技术中心
社招全职3年以上技术类-算法地点:北京状态:招聘
工作描述
任职要求 学历与专业背景: 3年以上工作经验,硕士及以上学历,中大厂AI Infra从事经验 能力要求: 1. 有AI推理加速框架/引擎研发经验者优先,如参与过vLLM、TensorRT-LLM、Triton、OneFlow等相关框架的开发与优化。 2. 熟悉具身智能、交互式世界模型、机器人AI相关业务场景,了解实时推理、多模态融合推理、连续状态推理的技术难点者优先。 3. 具备分布式推理、多机多卡协同推理研发经验,熟悉NCCL、MPI等通信机制,能解决分布式场景下的通信瓶颈与调度问题。 4. 具备AI编译器、中间表示(IR)优化相关经验者优先,熟悉TVM、MLIR等编译工具链优先。 5. 具备良好的系统设计能力与问题排查能力,能独立解决复杂的性能瓶颈、内存溢出、硬件适配等工程问题。 6. 具备强烈的技术钻研精神,对AI基础设施、前沿AI场景落地有浓厚兴趣,善于攻克技术难题。 7. 良好的团队协作与沟通能力,能跨团队协同推进项目,清晰输出技术方案与成果。 8. 具备优秀的文档撰写与技术分享能力,注重代码质量与工程规范。 工作职责 我们是谁? 作为中国领先的数字地图内容及导航服务提供商,高德地图日均服务数亿用户出行决策,每日处理超百亿级位置数据。视觉技术中…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
vLLM+
https://www.newline.co/@zaoyang/ultimate-guide-to-vllm--aad8b65d
vLLM is a framework designed to make large language models faster, more efficient, and better suited for production environments.
https://www.youtube.com/watch?v=Ju2FrqIrdx0
vLLM is a cutting-edge serving engine designed for large language models (LLMs), offering unparalleled performance and efficiency for AI-driven applications.
TensorRT+
https://docs.nvidia.com/deeplearning/tensorrt/latest/getting-started/quick-start-guide.html
This TensorRT Quick Start Guide is a starting point for developers who want to try out the TensorRT SDK; specifically, it demonstrates how to quickly construct an application to run inference on a TensorRT engine.
Triton Inference Server+
https://docs.nvidia.com/deeplearning/triton-inference-server/user-guide/docs/index.html
Triton Inference Server is an open source inference serving software that streamlines AI inferencing.
NCCL+
https://developer.nvidia.com/nccl
The NVIDIA Collective Communication Library (NCCL) implements multi-GPU and multi-node communication primitives optimized for NVIDIA GPUs and networking.
Message Passing Interface+
https://www.youtube.com/watch?v=7huftuXExV0
Parallel programming and MPI are crucial tools for achieving high performance computing.
[英文] 📺Basics of the Message Passing Interface (MPI) to program distributed memory parallel computers
https://www.youtube.com/watch?v=tm8M5H1OZmw
The Message Passing Interface (MPI) is a widely used standard to program distributed message parallel computers.
系统设计+
https://roadmap.sh/system-design
Everything you need to know about designing large scale systems.
https://www.youtube.com/watch?v=F2FmTdLtb_4
This complete system design tutorial covers scalability, reliability, data handling, and high-level architecture with clear explanations, real-world examples, and practical strategies.
自动驾驶+
https://www.youtube.com/watch?v=_q4WUxgwDeg&list=PL05umP7R6ij321zzKXK6XCQXAaaYjQbzr
Lecture: Self-Driving Cars (Prof. Andreas Geiger, University of Tübingen)
https://www.youtube.com/watch?v=NkI9ia2cLhc&list=PLB0Tybl0UNfYoJE7ZwsBQoDIG4YN9ptyY
You will learn to make a self-driving car simulation by implementing every component one by one. I will teach you how to implement the car driving mechanics, how to define the environment, how to simulate some sensors, how to detect collisions and how to make the car control itself using a neural network.
还有更多 •••
相关职位
社招微信AI助手技术
1.熟练掌握 C/C++、Python语言,有计算机体系结构背景或软件开发背景,熟悉系统性能调优的方式; 2.具备基础的GPU编程能力,包括但不限于Cuda、OpenCL;熟悉至少一种GPU加速库,
更新于 2026-09-15北京
社招1年以上搜一搜技术
1.岗位要求:; 2.熟悉AI基础硬件设置,有真实的大规模推理系统的设计开发部署经验; 3.熟悉各种主流LLM/VLM的模型结构,具有 vllm/sglang/TRT-llm等推理引擎优化实践经验
更新于 2026-06-11北京
社招1年以上搜一搜技术
1.熟悉AI基础硬件设置,有真实的大规模推理系统的设计开发部署经验; 2.熟悉各种主流LLM/VLM的模型结构,具有 vllm/sglang/TRT-llm等推理引擎优化实践经验; 3.熟悉LLM
更新于 2026-06-11北京