
商汤大模型训练实习生
实习兼职互联网 / 电子 / 网游地点:深圳状态:招聘
工作描述
任职要求 1、扎实的机器学习 / 深度学习基础 2、较强的 Coding 能力 3、较强的实验设计与问题分析能力 4、对大模型训练 / 强化学习有强烈兴趣 5、实习时间稳定 加分项: 有 LLM / VLM / MLLM 训练经验,或相关方向 paper / 开源项目; 有 SFT / RLHF / RLAIF / DPO / PPO / GRPO 实践经验; 有 Reward Model / Preference Model 训练或 Reward Engineering 经验; 有 VLM-as-a…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
还有更多 •••
相关职位

实习算法
1.本科及以上学历、计算机、软件工程等相关专业优先; 2.有扎实的计算机科学知识,掌握Pytorch,具备良好的编程能力和代码风格。 3. 对AI大模型相关核心技术感兴趣, 对megatron dee
更新于 2026-07-09上海

实习算法
职位要求 任职要求 1. 计算机科学、人工智能、机器学习等相关专业在读硕士或博士(优秀本科生可考虑) 2. 具备扎实的深度学习基础,了解 LLM / VLM 等模型结构 3. 对分布式训练 / 推理
更新于 2026-07-09北京|上海|成都
实习
工作内容: 1、参与分布式推理系统建设,提供行业领先的LLM/多模态模型解决方案; 2、针对大语言模型(LLM)等场景,开展端到端推理性能优化,重点降低推理延迟、提升吞吐量; 3、针对大语言模型(LL
北京