阿里巴巴阿里国际-高级AI算法工程师-杭州
社招全职1年以上阿里集团地点:杭州状态:招聘
工作描述
任职要求 模型与后训练:精通 Transformer/LLM 架构;具备 SFT/DPO/RLHF 全链路实操经验;有 Agentic RL 训练经验者优先。 Agent 与系统:能独立完成任务拆解与多 Agent 协作编排;熟练工程化落地 RAG、Memory 及 Tool-Use(含 MCP 等标准协议)。 数据工程:具备 Data-centric 思维;精通高质量后训练数据挖掘与构造;有合成数据(Synthetic Data)与动作轨迹(Trajectory)构建经验者优先。 评测闭环:能搭建包含 LLM-as-judge、离线评测、A/B 测试的完整评估体系;具备复杂多步任务的量化评估与问题定位能力。 工…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
RAG+
https://www.youtube.com/watch?v=sVcwVQRHIc8
Learn how to implement RAG (Retrieval Augmented Generation) from scratch, straight from a LangChain software engineer.
MCP+
https://www.youtube.com/watch?v=eur8dUO9mvE
Unlock the secrets of MCP! 🚀 Dive into the world of Model Context Protocol and learn how to seamlessly connect AI agents to databases, APIs, and more. Roy Derks breaks down its components, from hosts to servers, and showcases real-world applications. Gain the knowledge to revolutionize your AI projects!
https://www.youtube.com/watch?v=L94WBLL0KjY
Let's talk about MCP or the Model Context Protocol.
数据挖掘+
https://www.youtube.com/watch?v=-bSkREem8dM
Database vs Data Warehouse vs Data Lake
https://www.youtube.com/watch?v=7rs0i-9nOjo
还有更多 •••
相关职位
社招2年以上云智能集团
1. 计算机、人工智能等相关专业,2年以上搜索、推荐、NLP、大模型等领域算法经验。 2. 熟练掌握机器学习/深度学习算法,熟悉deepspeed、megatron等大模型训练框架,在搜推/NLP/大
更新于 2026-01-22北京|杭州
社招算法
1. 熟悉 LLM / MLLM / AIGC 原理,有大模型后训练、微调、蒸馏、RAG、Agent 调优等完整项目经验,能独立负责复杂算法系统的端到端建设 2. 具备扎实的机器学习、统计学和数据挖掘
更新于 2026-06-01深圳
社招1-10年SOFTWARE
3年以上语音算法或端侧 AI 算法研发经验,有终端语音交互落地经验; 扎实的语音信号处理 / 语音唤醒算法基础,熟悉语音降噪/回声消除/语音唤醒/声纹识别等相关算法原理; 精通 C/C++ 与 Jav
更新于 2026-08-23北京