蚂蚁金服【蚂蚁星】财富AI Lab-大模型训练与应用-27届
校招全职2027届蚂蚁星- Plan A人才计划地点:北京 | 上海 | 杭州状态:招聘
工作描述
任职要求 1. 具备扎实的大模型后训练经验,熟悉 SFT、DPO、PPO、GRPO、RLHF/RLAIF、Reward Model、Preference Data 等相关技术原理和实践方法; 2. 具备较强的算法研发能力和实验设计能力,能够基于业务问题拆解模型能力短板,设计数据、训练、评测和分析方案; 3. 熟悉主流大模型训练和微调框架,如 Megatron-LM、DeepSpeed、FSDP、LLaMA-Factory、ms-swift、Transformers 等,有大规模训练或微调经验者优先; 4. 熟悉大模型数据工程,具备大规模语料处理、指令数据构建、合成数据生成、数据质量评估、偏好数据构建等经验; 5. 具备扎实的数学和机器学习基础,熟悉线性代数、概率统计、优化算法、深度学习和强化学习相关知识; 6. 对 LLM、Agent、Tool Use、RAG、多轮对话、复杂推理等方向有深入理解,能够将前沿技术转化为可落地的产品能力。 【加分项(模型×Agent联合优化成果)】 1.在模型与Agent的交叉领域有过充分实践优先: 2.顶会论文:NeurIPS/ICML/ICLR/ACL等与Agen…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
Megatron+
https://www.youtube.com/watch?v=hc0u4avAkuM
LLaMA-Factory+
https://llamafactory.readthedocs.io/en/latest/
LLaMA Factory is an easy-to-use and efficient platform for training and fine-tuning large language models.
Swift+
[英文] A Swift Tour
https://docs.swift.org/swift-book/documentation/the-swift-programming-language/guidedtour/
Explore the features and syntax of Swift.
https://www.hackingwithswift.com/learn
Free Swift and iOS tutorials
https://www.youtube.com/watch?v=8Xg7E9shq0U
Learn the Swift programming language in this full tutorial for beginners.
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
还有更多 •••
相关职位
校招2027届蚂蚁星
1. 扎实的大模型算法基础,理解 Transformer、LLM 训练、后训练、推理机制和对齐技术; 2. 熟悉 Agent 技术范式,如 ReAct、Plan-and-Execute、Reflexi
更新于 2026-05-14北京|上海|杭州
校招2027届蚂蚁星
1. 计算机、人工智能、软件工程、机器学习等相关专业本科及以上学历,具备扎实的工程基础; 2. 熟悉大模型基本原理,理解 Transformer、LLM 推理机制、KV Cache、Attention
更新于 2026-05-14北京|上海|杭州
实习蚂蚁星- Pla
1. 计算机科学、人工智能、数学等相关专业硕士及以上学历,博士优先; 2. 深入掌握Transformer/BERT/GPT等架构,有1个以上千亿参数大模型实战经验(训练/推理/优化全流程); 3.
更新于 2025-07-25北京|上海|杭州