美团大模型算法研究员-复杂推理/RL方向
社招全职1年以上核心本地商业-基础研发平台地点:北京 | 上海状态:招聘
工作描述
任职要求 1. 数学、物理、计算机和机器学习等相关专业 2. 具备post-training或强化学习相关经验 3. 熟悉至少一种深度学习框架和强化学习训练框架,具备良好的算法工程结合能力 4. 有AGI信仰,对大模型相关技术有浓厚兴趣,具备强烈的进取心、求知欲,热衷于追求行业前沿的技术创新。 5. 在国内外知名大模型团队有研究和实践经验,…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
还有更多 •••
相关职位
社招算法开发岗
1.扎实的机器学习、NLP、RL基础和出色的创新能力,在ACL/EMNLP/NAACL/NeurIPS/ICML/ICLR等顶级会议上发表论文者优先; 2.在预训练、后训练、强化学习方向有深刻研究;
更新于 2026-03-25北京

校招ai 算法类
1、计算机区相关专业硕士及以上学历,博士优先; 2、熟练掌握SFT, RLHF,Agentic RL等post-training优化方法,熟悉Slime, VeRL等Agentic RL模型训练框架,
杭州

校招研发
1.本科及以上学历,计算机科学与技术、人工智能、软件工程、电子信息、数学等相关专业。 2.具备扎实的机器学习、深度学习及自然语言处理基础,熟悉 Python 编程,熟悉 PyTorch 等主流深度学习
更新于 2026-07-06深圳