
商汤大模型强化学习系统研究员
校招全职算法地点:北京 | 深圳 | 上海 | 成都状态:招聘
工作描述
任职要求 1、2027届本科及以上学历,计算机、软件工程、人工智能等相关专业; 2、 具备扎实的深度学习基础,了解 LLM / VLM 等大模型结构与训练流程; 3、 对分布式训练、推理加速、强化学习系统或大模型系统优化有浓厚兴趣,并具备一定实践经验; 4、熟悉 PyTorch,了解或使用过 …
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
还有更多 •••
相关职位

校招算法
1、2027届本科及以上学历,计算机、软件工程、人工智能等相关专业; 2、 具备扎实的深度学习基础,了解 LLM / VLM 等大模型结构与训练流程; 3、 对分布式训练、推理加速、强化学习系统或大模
更新于 2026-07-07北京|深圳|上海
实习阿里巴巴研究型实
1. 在读博士研究生,计算机相关专业; 2.有大模型后训练或者强化学习相关经验和工作背景; 3.有A类会议或者期刊论文发表经历; 4.有较强的代码能力,熟练掌握Python编程。 工作职责 我们正在
更新于 2026-05-06北京|杭州
实习阿里巴巴研究型实
1. 在读博士研究生,计算机相关专业; 2.有大模型后训练或者强化学习相关经验和工作背景; 3.有A类会议或者期刊论文发表经历; 4.有较强的代码能力,熟练掌握Python编程。 工作职责 我们正在
更新于 2026-03-17北京|杭州