菜鸟菜鸟-具身智能强化学习算法工程师-杭州
社招全职3年以上技术类-算法地点:杭州状态:招聘
工作描述
任职要求 1. 计算机/自动化/机器人相关专业,扎实的 Python 工程能力。 2. 系统掌握强化学习原理,熟悉各类 On/Off-Policy 等强化学习算法及其调参与稳定性问题。 3. 了解模仿学习与 VLA 训练范式,理解预训练与后训练的分工与衔接。 4. 熟悉至少一种仿真/训练平台(Isaac Lab/Isaac Gym/MuJoCo),具备真机或仿真 RL 落地经验。 加分项: 1. 有大模型/VLA 后训练经验(RLHF/RLAIF、偏好优化、奖励建模)。 2. 有 HIL-SERL、RE…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
还有更多 •••
相关职位
校招
1.硕士及以上学历在读,计算机、人工智能、机器人、自动化等相关专业,2026 年及以后毕业优先 2.深刻理解强化学习核心算法(PPO, SAC 等),同时熟悉具身操作大模型(VA/VLA/WAM)的训
更新于 2026-06-02深圳
社招3年以上无人机业务部
学历背景: 计算机、自动化、机器人等相关专业毕业,对人工智能、计算机图形学、机器人学有深入的理论理解。 编程能力: 精通 Python 及 PyTorch 等主流框架,具备扎实的算法实现能力;熟悉 C
更新于 2026-06-17北京|深圳
校招人工智能
1、熟练掌握大语言模型及多模态大模型的训练和优化方法,精通主流深度学习框架如 PyTorch 或 TensorFlow,熟练掌握 Python 或 C++ 至少一种编程语言; 2、具备过硬的工程能力,
杭州