阿里巴巴日常实习生-大模型后训练/Agentic算法(行为风控方向)
实习兼职阿里巴巴日常实习生地点:北京状态:招聘
工作描述
任职要求 1、硕士研究生及以上在读,计算机、人工智能、软件工程、信息安全、统计、数学等相关专业; 2、具备大模型、强化学习、Agent相关科研、项目或实习经验,复合能力较强者优先; 3、了解业务风控相关场景者优先,包括作弊、欺诈、账号安全、恶意行为等;有日志、行为序列、图数据处理经验者优先; 4、具备扎实的编程基础,熟练掌握 Python,熟悉 PyTorch 等深度学习框架; 5、具备良好的分析问题和动手能力,能够在指导下将业务问题抽象为可实验、可建模的技术问题; 6、具备良好的沟通能力和团队合…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
还有更多 •••
相关职位
实习阿里巴巴日常实习
1. 对大模型后训练有深入理解和数据直觉,熟悉常见的数据构造策略和 RL 算法,具备算法设计到代码落地的完整经验。 2. 具有较强的代码工程能力,精通 Python 以及 Pytorch 等深度学习框
更新于 2026-08-06北京|杭州|上海
实习阿里巴巴日常实习
1、硕士及以上学历在读,计算机、人工智能、网络安全等相关专业优先。 2、具备扎实的深度学习与自然语言处理基础,熟悉深度学习框架(PyTorch),熟练掌握 Prompt Engineering、模型微
更新于 2026-07-10杭州
实习阿里巴巴日常实习
1. 计算机、人工智能、数据科学、自动化等相关专业,本科及以上学历; 2. 对大语言模型(LLM)及 AI 应用有浓厚兴趣,有相关课程学习、科研项目或竞赛经验(涉及 NLP、大模型微调、RAG 检索增
更新于 2026-06-29北京