阿里巴巴达摩院-医疗AI Agent 算法工程师-杭州
社招全职1年以上地点:杭州状态:招聘♡ 收藏
工作描述
任职要求 ● 计算机、人工智能、生物医学工程、数学等相关专业硕士及以上。 ● 精通 Python 与 PyTorch,具备分布式训练与推理优化的工程能力。 ● 在以下至少一个方向具备扎实的理论与实战经验: ● Agent 系统:任务规划、工具调用、多轮决策与编排框架;熟悉 ReAct、Plan-and-Execute、Reflexion 等主流范式。 ● Agentic RL:PPO / GRPO / DPO / RLVR、奖励建模、过程奖励(PRM)、Self-play、多轮 rollout 训练。 ● 大模型训练与对齐全流程:SFT → RLHF / RL,熟悉主流架构、推理机制与 Test-time Scaling / Reasoning Model。 ● RAG 与对话系统:检索、召回、重排、引用生成、多轮对话状态管理。 …
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
React+
[英文] Quick Start - React
https://react.dev/learn
This page will give you an introduction to 80% of the React concepts that you will use on a daily basis.
https://www.youtube.com/watch?v=SqcY0GlETPk
Master React 18 with TypeScript! ⚛️ Build amazing front-end apps with this beginner-friendly tutorial.
https://www.youtube.com/watch?v=x4rFhThSX04
Learn modern React basics in the most interactive, hands-on way possible in the full course for beginners.
GRPO+
https://cameronrwolfe.substack.com/p/grpo
Most early work on RL for LLMs used Proximal Policy Optimization (PPO) as the default RL optimizer, but recent reasoning research relies upon Group Relative Policy Optimization (GRPO).
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
还有更多 •••
相关职位

社招1年以上技术类-算法
1、计算机科学、人工智能、电子工程、生物医学工程等相关专业的硕士或博士毕业,在视觉生成或者医疗影像AI上有实际的研发经历,精通深度学习、机器学习、计算机视觉及医学影像处理,工业界相关研发工作经验不少于
更新于 2026-04-01北京|杭州

社招5年以上产品类-商业型
1. 8年以上医疗AI/数字化医疗产品经验,主导过至少1款成功落地的医疗AI产品。 2. 熟悉医院诊疗流程(如信息科、影像科、相关专科工作流等)及医疗数据应用规范。 3. 强市场洞察:能通过政策解读、
更新于 2026-04-07北京|杭州
社招5年以上技术类-开发
● 硕士及以上学历,生物医学工程、医疗器械、机械、电子、电气、自动化、机器人、软件工程或相关专业。 ● 5 年以上医疗器械注册、法规、质量体系、硬件产品、系统工程或项目管理经验。 ● 必须具备完整三类
更新于 2026-06-09杭州