小米足式机器人运动控制算法工程师实习生(强化学习)
实习兼职地点:北京状态:招聘
工作描述
任职要求 1. 硕士及以上学历,机器人、计算机、机械工程、人工智能、应用数学等专业,数学、英语能力扎实,具有较强的学习与研究能力; 2. 掌握主流的强化学习算法,如:PPO、DQN、DDPG、SAC等; 3. 掌握机器人学习中的广泛使用的训练方法和模型架构,如:教师学生模型(Teacher-Student Network),课程学习(Curriculum Lear…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
C+++
https://www.learncpp.com/
LearnCpp.com is a free website devoted to teaching you how to program in modern C++.
https://www.youtube.com/watch?v=ZzaPdXTrSb8
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
还有更多 •••
相关职位
社招3年以上智能与信息技术
1、机器人、控制、自动化、机械、计算机等相关专业,本科及以上学历。 2、3年以上强化学习或机器人运动控制相关经验;具备足式/双足机器人真机落地经验者优先。 3、熟悉主流RL算法与实现(PPO/SAC/
北京
社招
1. 熟练掌握maya绑定流程以及相关工具,熟悉游戏开发相关rigging流程,有至少1款A级品质游戏开发经验,如只有影视行业资深绑定经验,对游戏开发有强烈愿望,愿意学习游戏开发流程者亦可; 2. 有
更新于 2025-11-26上海
校招
1.硕士及以上学历在读,计算机、人工智能、机器人、自动化、控制工程等相关专业,2027年及以后毕业优先 2.深刻理解强化学习核心算法(PPO、SAC等),熟悉 Actor-Critic 架构及其在连续
更新于 2026-06-17深圳