安克创新具身智能-强化学习(灵巧操作方向) 实习生
校招全职地点:深圳状态:招聘
工作描述
任职要求 1.硕士及以上学历在读,计算机、人工智能、机器人、自动化等相关专业,2026 年及以后毕业优先 2.深刻理解强化学习核心算法(PPO, SAC 等),同时熟悉具身操作大模型(VA/VLA/WAM)的训练逻辑 3.具备扎实的机器人运动学、动力学基础,能够处理真机实验中的延迟、噪声及硬件非线性特性 4.精通 Python 与 PyTorch,熟悉主流 RL 框架,具备良好的分布式训练与真机部署工程经验 5.了解以下至少一个方向的核心技术: Offline-to-online 真机RL算法 去噪模型(Flow matching/Diffusion)RL算法 VLA / 多模态大模型 机器人学习或具身智能基础方法 6.具备较强的论文阅读与复现能力,能够…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
还有更多 •••
相关职位
社招3年以上技术类-算法
1. 计算机/自动化/机器人相关专业,扎实的 Python 工程能力。 2. 系统掌握强化学习原理,熟悉各类 On/Off-Policy 等强化学习算法及其调参与稳定性问题。 3. 了解模仿学习与 V
更新于 2026-08-07杭州
社招3年以上无人机业务部
学历背景: 计算机、自动化、机器人等相关专业毕业,对人工智能、计算机图形学、机器人学有深入的理论理解。 编程能力: 精通 Python 及 PyTorch 等主流框架,具备扎实的算法实现能力;熟悉 C
更新于 2026-06-17北京|深圳
校招人工智能
1、熟练掌握大语言模型及多模态大模型的训练和优化方法,精通主流深度学习框架如 PyTorch 或 TensorFlow,熟练掌握 Python 或 C++ 至少一种编程语言; 2、具备过硬的工程能力,
杭州