阿里巴巴研究型实习生-Agentic Search-Accio
实习兼职阿里巴巴日常实习生地点:杭州状态:招聘
工作描述
任职要求 1、2027年11月及之后毕业,硕士及以上学历,计算机、人工智能、信息科学、数学、软件工程等相关专业; 2、具有研究经验:大模型 post-training / 强化学习方向,具有一篇及以上一作顶会论文(NeurIPS / ICML / ICLR / ACL 等); 3、技术栈:对 post-training / RL 有实际研究经验,熟悉主流训练框架; 4、能力:良好的论文阅读与工程实现能力,能快速跟踪前沿并独立推进研究。 加分项:有复杂推理、AI搜索、RAG 等方向的研究经验或顶会论文;具备跨境电商或 B2B 贸易场景经验。 工作职责 团队介绍: Accio 是阿…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
NeurIPS+
https://neurips.cc/
还有更多 •••
相关职位
实习阿里巴巴研究型实
1、2027年11月及之后毕业,硕士及以上学历,计算机、人工智能、信息科学、数学、软件工程等相关专业; 2、具有研究经验:大模型 post-training / 强化学习方向,具有一篇及以上一作顶会论
更新于 2026-08-21杭州
实习阿里巴巴研究型实
1. 计算机、人工智能、统计等相关专业在读博士。 2. 扎实的算法基础,熟练使用 Python 与 PyTorch,具备独立完成实验设计、执行与分析的能力。 3. 熟悉 Agent / Tool-Us
更新于 2026-08-07北京|杭州
实习阿里巴巴研究型实
1. 背景: 计算机、数学或相关专业在读(硕士/博士优先),有agent实践经验。 2. 技术栈: 熟悉 PyTorch,有agent 框架实践经验。 3. 良好的论文阅读与工程实现能力,希望在顶会发
更新于 2026-07-07杭州