京东大模型算法工程师(AI Infra 高阶)
社招全职算法开发岗地点:上海状态:招聘
工作描述
任职要求 1. 具备大规模语言模型与视觉语言模型全流程训练经验,能设计并参与预训练、中训练及后训练阶段,保障模型能力持续提升; 2. 拥有万卡级别集群下AI基础设施实践经验,熟悉从数据到训练再到推理的完整链路,能基于Megatron、vLLM等框架优化分布式训练与推理性能; 3. 能够定位大规模训练故障或性能瓶颈的底层成因,制定有效对策,并具备结合前沿论文与工程实践在数据筛选、强化学习对齐或算子优化等方面提出创新性解决方案的能力; 4. 致力于在大模型基础架构领域追求技术领先,长期投入于AI基础设施的突破,并在高强度技术攻坚中坚持高标准,主动承担难点任务,追求持续技术突破; 5. 符合京东价值观:客户为先、创新、拼搏、担当、感恩、诚信。 工作职责 我们的核心使命 (Core Mission):我们专注于大语言模型(LLM)的全栈技术突破与前沿探索,拒绝平庸的复刻,追求极致的创新: 大模型基座研发:从 Pre-training 到 Post-training,探索千亿/…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
Megatron+
https://www.youtube.com/watch?v=hc0u4avAkuM
vLLM+
https://www.newline.co/@zaoyang/ultimate-guide-to-vllm--aad8b65d
vLLM is a framework designed to make large language models faster, more efficient, and better suited for production environments.
https://www.youtube.com/watch?v=Ju2FrqIrdx0
vLLM is a cutting-edge serving engine designed for large language models (LLMs), offering unparalleled performance and efficiency for AI-driven applications.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
智能体+
https://learn.microsoft.com/en-us/shows/ai-agents-for-beginners/
In this 10-lesson course we take you from concept to code while covering the fundamentals of building AI agents.
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
还有更多 •••
相关职位
社招算法开发岗
1. 具备大规模语言模型与视觉语言模型全流程训练经验,能设计并参与预训练、中训练及后训练阶段,保障模型能力持续提升; 2. 拥有万卡级别集群下AI基础设施实践经验,熟悉从数据到训练再到推理的完整链路,
更新于 2026-08-03上海
校招研发类
24-26届博士,计算机/AI相关等相关专业博士,有AI算法基础和实战能力,能快速参与AI中台开发和优化工作。 工作职责 1、业务介绍: 负责顺丰统一AI和智能体基础设施建设,打造行业领先的AI中台
更新于 2026-07-08深圳
校招A243713
1、2027届获得硕士及以上学位,计算机、人工智能、自动化、数学相关专业优先; 2、拥有AI coding相关研究经验,或拥有Agent优化、大模型训练相关实操经验者优先; 3、拥有AI领域的研究经历
更新于 2026-07-29深圳|上海