美团【北斗】大模型算法研究员(Agent/RL/推理)
校招全职核心本地商业-业务研发平台地点:北京状态:招聘
工作描述
任职要求 【任职资格】 必备条件: 1.2027届计算机、数学、统计等相关专业在读或应届,本科及以上,博士/硕士优先 2.扎实的机器学习与深度学习基础,熟悉Transformer架构及其变体,具备独立阅读和复现顶会论文的能力 3.熟练掌握Python及PyTorch/JAX等主流框架,具备清晰的代码工程意识 4.对大模型的训练流程(预训练/后训练)或Agent构建有系统性理解,具备独立完成端到端实验的能力 5.具备RLHF/DPO/GRPO或其他对齐算法的实际训练与调优,对相关数据构建有深度认知 加分项: 1.熟悉ClaudeCode、OpenClaw、Hermes等开源Harness的设计和实现 2.在NeurIPS/ICML/ICLR/ACL/EMNLP等顶会发表过论文(含在投),或有被广泛引用的开源项目 3.有Agent系统(如ReAct/Toolformer/CodeAct类)的研究或工程经验,理解Agent失败模式与评估瓶颈 4.参与过千卡以上规模分布式训练,或对推理优化(量化、投机解码等)有动手经验 5.ACM-ICPC/Kaggle/算…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
JAX+
https://docs.jax.dev/en/latest/notebooks/thinking_in_jax.html
JAX is a library for array-oriented numerical computation, with automatic differentiation and JIT compilation to enable high-performance machine learning research.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
校招核心本地商业-基
【岗位要求】 1. 2027届硕士及以上学历,计算机、人工智能、数学等相关专业。 2. 扎实的深度学习理论基础,精通PyTorch等主流框架,具备独立复现和改进前沿模型的能力。 3. 在NLP/多模态
更新于 2026-06-03北京|上海
实习核心本地商业-业
海内外高校在校本科生(大三及以上)、硕士生及博士生,且以下条件至少满足一项: 1.超级学霸:专业成绩排名前1%。 2.学术达人:在顶级期刊或学术会议上以第一作者身份发表论文(或导师一作,自己为二作)。
更新于 2026-02-06北京
校招核心本地商业-业
【任职资格】 必要条件: 1.2027届本科及以上学历,计算机、人工智能等相关专业; 2.在大模型后训练等方面有深入实践,具备较强的动手能力; 3.扎实的深度学习和计算机理论基础,精通主流深度学习框架
更新于 2026-06-03北京