蚂蚁金服【蚂蚁星】财富AI Lab-大模型Agent算法-27届
校招全职2027届蚂蚁星- Plan A人才计划地点:北京 | 上海 | 杭州状态:招聘
工作描述
任职要求 1. 扎实的大模型算法基础,理解 Transformer、LLM 训练、后训练、推理机制和对齐技术; 2. 熟悉 Agent 技术范式,如 ReAct、Plan-and-Execute、Reflexion、Tree-of-Thought、Function Calling、Multi-Agent 等; 3. 熟悉 Tool Use 相关技术,理解工具选择、参数生成、工具结果解析、多工具协同和失败恢复等问题; 4. 熟悉 SFT、DPO、PPO、GRPO、RLHF/RLAIF、Reward Model、Preference Learning 等方法; 5. 具备较强工程能力,熟练使用 Python,熟悉 PyTorch、Transformers、vLLM、Ray、DeepSpeed、Megatron-LM、FSDP、OpenRLHF、verl、slime 等框架中的一种或多种; 6. 具备良好的任务抽象和实验分析能力,能够将复杂任务拆解为可执行、可评测、可训练的 Agent 任务; 7. 对大模型 Agent 方向有持续热情,关注前沿研究,对 Agent 失败模式和真实系统落地挑战有深入思考。 【…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
React+
[英文] Quick Start - React
https://react.dev/learn
This page will give you an introduction to 80% of the React concepts that you will use on a daily basis.
https://www.youtube.com/watch?v=SqcY0GlETPk
Master React 18 with TypeScript! ⚛️ Build amazing front-end apps with this beginner-friendly tutorial.
https://www.youtube.com/watch?v=x4rFhThSX04
Learn modern React basics in the most interactive, hands-on way possible in the full course for beginners.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
还有更多 •••
相关职位
校招2027届蚂蚁星
1. 计算机、人工智能、软件工程、机器学习等相关专业本科及以上学历,具备扎实的工程基础; 2. 熟悉大模型基本原理,理解 Transformer、LLM 推理机制、KV Cache、Attention
更新于 2026-05-14北京|上海|杭州
校招2027届蚂蚁星
1. 具备扎实的大模型后训练经验,熟悉 SFT、DPO、PPO、GRPO、RLHF/RLAIF、Reward Model、Preference Data 等相关技术原理和实践方法; 2. 具备较强的算
更新于 2026-05-14北京|上海|杭州
实习蚂蚁星- Pla
1. 计算机科学、人工智能、数学等相关专业硕士及以上学历,博士优先; 2. 深入掌握Transformer/BERT/GPT等架构,有1个以上千亿参数大模型实战经验(训练/推理/优化全流程); 3.
更新于 2025-07-25北京|上海|杭州