哔哩哔哩Omni理解模型算法实习生
实习兼职技术类地点:上海状态:招聘
工作描述
工作职责: 1. 理解模型训练:参与多模态大模型(文本、图像、视频等)的算法研究与工程实现,重点攻关模型在复杂视觉场景下的视频理解能力。 2. SFT专项迭代:跟踪学术界与工业界前沿进展,方向包括但不限于鲁棒视觉表征与视觉特征编解码、多模态模型的RLVR 与 RLHF、Synthetic Data、视觉审美与视觉品质评估等,持续打磨模型对视觉世界的"体感",对构图、布局、风格、细节的直觉判断力。 3. 实验与复盘:设计与开发高效的分布式训练与数据系统,优化大规模多模态模型的训练流程与数据处理链路,支撑海量多模态数据的高效存储、检索、清洗与迭代,保障上述方向的快速落地 工作要求: omni理解模型算法实习 职位描述 工作职责 1. 理解模型训练:参与…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
NeurIPS+
https://neurips.cc/
还有更多 •••
相关职位
社招程序&技术类
1、计算机科学、人工智能、电子工程等相关博士学历; 2、具备大模型(LLM 或多模态)训练经验,深入理解 Transformer 架构及分布式训练框架(Megatron-LM, DeepSpeed,T
上海
社招3年以上技术类-算法
1. 计算机科学、电子工程、人工智能/机器学习或相关领域硕士及以上学历,具备2年以上多模态 AI 或语音处理相关行业经验。 2. 精通多模态学习(视觉-语言、音频-视觉或全模态模型),熟练使用深度学习
更新于 2026-07-28杭州
社招1年以上技术类-算法
1. 计算机科学、人工智能、机器学习、语音处理或相关领域硕士及以上学历。 2. 具备大模型、对话系统、多模态模型或语音交互方向研发经验,对实时对话式 AI 产品有较深理解,具备较强的工程落地与问题抽象
更新于 2026-06-12北京|杭州