理想汽车【基座模型】高级算法工程师(智能体强化学习方向)
社招全职汽车研发地点:北京状态:招聘
工作描述
任职要求 任职要求 计算机科学、数学、人工智能、统计学等相关专业,硕士及以上学历。 具备深厚的强化学习(RL)理论功底,熟悉传统 RL 与 LLM 结合的最新进展,对 Agent-RL、Search-RL、Online RL 或决策智能体有深入理解。 精通 PyTorch,拥有出色的工程素养。能够高效处理大规模训练数据,具备在大规模 GPU 集群上进行模型分布式训练与算子…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
智能体+
https://learn.microsoft.com/en-us/shows/ai-agents-for-beginners/
In this 10-lesson course we take you from concept to code while covering the fundamentals of building AI agents.
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
还有更多 •••
相关职位
社招智能与信息技术
1.硕士及以上学历,自动化、计算机等相关专业; 2.必须有视觉VIT基座优化经验或者VLM、VLA相关项目落地交付经验; 3.优秀的代码技术功底,熟练掌握CLIP编码器的From Scratch训练、
北京
社招3年以上智能与信息技术
任职要求: 本科及以上学历,3年及以上智能驾驶行车NOA测试经验 对所在片区(华中、华南、西南)的典型城市的道路结构及交通状况熟悉,当地候选人优先 候选人应具备客观的判断力、高度的组织能力、极强的细节
重庆
社招智能与信息技术
熟练掌握至少一门主流后端语言(Python / TypeScript / Go) 熟练使用常见中间件:Redis、Kafka、MySQL、MongoDB 熟悉工作流编排工具(如 Temporal、Ai
北京