阿里巴巴研究型实习生-智能算法产品事业部-推荐Agent算法工程师
实习兼职阿里巴巴研究型实习生地点:北京状态:招聘
工作描述
任职要求 1. 计算机科学、软件工程、人工智能等相关专业在读硕士或博士生,具备硬核的理论基础; 2. 对推荐系统、LLM、Agent、强化学习等相关技术具有浓厚兴趣,具备天然的自驱力; 3. 具备扎实的编程功底和的丰富的工程动手经验; 4. 有以下经历者优先:1)参与过 LLM 预训练、SFT、RL、推理优化等实践过程并对原理有深入理解;2)有LLM推荐、Agent Memory等相关研究或实习经验;3)在…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
推荐系统+
[英文] Recommender Systems
https://www.d2l.ai/chapter_recommender-systems/index.html
Recommender systems are widely employed in industry and are ubiquitous in our daily lives.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
还有更多 •••
相关职位
实习阿里巴巴研究型实
1. 计算机科学、人工智能、自然语言处理、数据挖掘、机器学习等相关专业在读硕士或博士生,具备扎实的理论基础,有一定的科研能力,能够针对性的分析问题和解决问题; 2. 熟悉多模态表征学习及搜广推召回/排
更新于 2026-09-01北京|杭州
实习阿里巴巴研究型实
1.具备优秀的编码能力,扎实的数据结构和算法功底; 2.有一定的科研能力,能够针对性的解决问题; 3.优秀的分析问题和解决问题的能力,对解决具有挑战性问题充满激情,不轻易放弃; 4.对技术有热情,有自
更新于 2026-03-20北京
实习淘天集团研究型实
1.具备优秀的编码能力,扎实的数据结构和算法功底; 2.有一定的科研能力,能够针对性的解决问题; 3.优秀的分析问题和解决问题的能力,对解决具有挑战性问题充满激情,不轻易放弃; 4.对技术有热情,有自
更新于 2026-01-27北京