京东算法专家
社招全职算法开发岗地点:北京状态:招聘♡ 收藏
工作描述
任职要求 1.计算机、人工智能、数学、统计或相关专业硕士及以上学历; 2.熟悉强化学习、生成式建模(LLM/Diffusion/VAE/RLHF/Decision Transformer等)中的一种或多种; 3.熟练掌握Python,熟悉主流深度学习框架(PyTorch/TensorFlow); 4.具备广告推荐、自动出价、预算分配或多目标…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
还有更多 •••
相关职位

社招
1. 本科及以上学历,计算机相关专业。 2. 具备卓越的编码能力,扎实的数据结构和算法基础,擅长Python、Java、C/C++中至少一门语言。 3. 对推荐算法和机器学习有热情,乐于学习、思考和创
更新于 2024-08-09杭州

社招技术类
1、具备优秀的编码能力,扎实的数据结构和算法功底; 2、优秀的分析问题和解决问题的能力,对解决具有挑战性问题充满激情; 3、对技术有热情,有良好的沟通表达能力和团队精神; 4、熟悉机器学习、自然语言处
更新于 2026-06-18上海|北京|杭州
社招5年以上信息技术类
1、5年及以上工作经验,熟悉广告/推荐系统整体链路。 2、有广告出价/计费/混排等模块算法策略优化经验,有ROI2/全站推广告业务经验优先。 3、熟悉常用机器学习/深度学习/调控算法/运筹学算法。 4
更新于 2026-07-21上海|深圳|北京
