淘宝闪购淘宝闪购-高级营销算法专家-杭州/上海
社招全职3年以上技术类-算法地点:杭州 | 上海状态:招聘
工作描述
任职要求 1、计算机、运筹学、人工智能、应用数学等相关专业硕士及以上学历,在生成式、强化学习领域有深入研究,发表过顶会论文者优先。主导过生成式定价/广告/推荐系统完整项目者优先。 2、精通Transformer/Diffusion/Decision Transformer等架构,有将其应用于序列决策或组合优化的落地经验,熟练掌握强化学习框架,具备复杂奖励系统设计与多目标优化能力。具备SFT/DPO/GRP…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
推荐系统+
[英文] Recommender Systems
https://www.d2l.ai/chapter_recommender-systems/index.html
Recommender systems are widely employed in industry and are ubiquitous in our daily lives.
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
还有更多 •••
相关职位

社招3年以上技术类-算法
1. 计算机、运筹学、应用数学、统计学或相关专业,有算法领域的工作经验; 2. 熟悉机器学习/深度学习/数据挖掘/运筹优化中至少一个领域的原理与算法,并且能够熟练建模解决业务问题; 3. 在一个或多个
更新于 2026-04-09上海
社招3年以上技术类-算法
经验要求 国内外重点大学硕士及以上学历,2年以上互联网头部公司推荐 / 营销 / 大模型算法实战经验;有出行领域、本地生活或交易类业务背景者优先。 能力要求 专业技能 • 精通主流推荐 / 营销算法
更新于 2026-06-05北京

社招3年以上技术类-算法
经验要求 国内外重点大学硕士及以上学历,2年以上互联网头部公司推荐 / 营销 / 大模型算法实战经验;有出行领域、本地生活或交易类业务背景者优先。 能力要求 专业技能 • 精通主流推荐 / 营销算法
更新于 2026-06-05北京