
商汤大模型后训练与智能体算法研发工程师
校招全职算法地点:北京 | 上海状态:招聘
工作描述
任职要求 1、2027 届硕士及以上学历,计算机、人工智能、软件工程、数据科学等相关专业优先; 2、熟练使用 Python,具备良好的代码能力和工程实现能力; 3、了解大模型基础概念,熟悉 SFT、RLHF/RLAIF、DPO、PPO、GRPO、Agent、Tool Use 等方向者优先; 4、对 Agentic RL、工具增强模型、多步推理或自动化任务解决有浓厚兴趣; 5、熟悉 Linux 开发环境,具备基础脚本开发、日志分析和问题定位能力; 6、具备较强的学习能力和抽象能力,…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
数据科学+
https://roadmap.sh/ai-data-scientist
Step by step roadmap guide to becoming an AI and Data Scientist
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
还有更多 •••
相关职位
实习阿里巴巴2027
1. 学术背景:计算机、人工智能、机器学习等相关方向的优秀硕士或博士毕业生,具备扎实的理论基础; 2. 代码能力:具备优秀的编码能力与数据结构基础,熟练掌握PyTorch,有大模型训练经验; 3. 问
更新于 2026-03-17北京|杭州
社招1年以上搜一搜技术
1.具备较强的动手能力;熟悉 Python ,具备扎实的系统编程功底和优秀的复杂系统 Debug 能力; 2.深入理解大模型分布式训练原理,具备 Megatron-LM、DeepSpeed 或 Py
更新于 2026-06-11广州
校招多模态大模型与应
1. 计算机科学、人工智能或机器学习等相关领域的本硕博在校生。 2. 深入理解大语言模型(LLM)底层架构,具备大模型预训练(Pre-training)、指令微调(SFT)或强化学习实战经验。 3.
更新于 2026-07-20北京