千问千问事业部-Omni全模态大模型算法专家-杭州/北京
社招全职1年以上技术类-算法地点:北京 | 杭州状态:招聘
工作描述
任职要求 1. 计算机科学、人工智能、机器学习、语音处理或相关领域硕士及以上学历。 2. 具备大模型、对话系统、多模态模型或语音交互方向研发经验,对实时对话式 AI 产品有较深理解,具备较强的工程落地与问题抽象能力。 3. 在以下一个或多个方向具备扎实经验:大模型后训练与对齐技术(如 SFT、RLHF、RLAIF、DPO、Online RL 等)、Omni/多模态模型训练优化、音视频通话场景下的数据与评测体系建设、聊天/问答/Agent 类场景的模型效果优化。 4. 熟悉大模型训练与推理相关技术栈,精通 Python,熟练使用 PyTorch 等主流框架,具备扎实的机器学习、深度学习与 Transformer/多模态模…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
学历+
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
社招程序&技术类
1、计算机科学、人工智能、电子工程等相关博士学历; 2、具备大模型(LLM 或多模态)训练经验,深入理解 Transformer 架构及分布式训练框架(Megatron-LM, DeepSpeed,T
上海
社招1年以上技术类-算法
1、计算机科学、人工智能、数学、电子信息工程或相关专业硕士及以上学历。 2、深入理解Transformer架构,熟悉SFT/RLHF/DPO/PPO/GRPO等算法原理及使用边界。 3、熟悉大模型训练
更新于 2026-07-13北京|杭州
实习阿里巴巴日常实习
1. 人工智能、计算机、声学等相关专业在读硕士或优秀本科生。 2. 熟悉 Python 及 PyTorch 框架,扎实掌握 Transformer 架构,对 LLM 原理有基本理解。 3. 对 Omn
更新于 2026-04-22北京|杭州