通义研究型实习生 - 下一代大模型评测方法与系统
实习兼职通义研究型实习生地点:北京 | 杭州 | 上海状态:招聘
工作描述
任职要求 1. 本科及以上学历,计算机、人工智能、软件工程、数学、自动化等相关专业优先; 2. 理论基础: 深入理解 Transformer 架构及大语言模型基础知识,熟悉模型评测方案(Evaluation)或具有后训练(Post-training,如 SFT、RLHF、DPO 等)经验; 3. 具备卓越的代码工程能力,精通 Python 编程及 PyTorch 深度学习框架;有模型训练、推理加速或自动化数据合成经验者优先; 4. 具备系统性研究思维,在国际顶级计算机会议/期刊(如 NeurIPS、ICML、ICLR、ACL、EMNLP 等)发表过一作论文,或在知名开源社区…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
还有更多 •••
相关职位
实习阿里巴巴研究型实
1. 本科及以上学历,计算机、人工智能、软件工程、数学、自动化等相关专业优先; 2. 理论基础: 深入理解 Transformer 架构及大语言模型基础知识,熟悉模型评测方案(Evaluation)或
更新于 2026-07-27北京|杭州|上海
实习阿里巴巴研究型实
1、国内、外高校博士,计算机/cv/nlp相关专业方向优先; 2、熟悉常用的大模型(LLMs)/多模态大模型(VLM)算法,具备极佳的代码工程能力,熟练使用c/c++/python等计算机语言,熟悉大
更新于 2026-03-19北京
实习阿里巴巴研究型实
1. 技术能力: a. 具备良好的C++/Python编程基础和代码实践能力。 b. 熟悉PyTorch。 c. 对底层技术有浓厚兴趣,渴望深入了解GPU架构与高性能计算。 2
更新于 2026-06-11北京|杭州