蚂蚁金服【Plan A】大模型动态评测与评测诊断方向(AGI蓝军)-27届
校招全职2027届蚂蚁星- Plan A人才计划地点:北京 | 上海 | 杭州状态:招聘
工作描述
任职要求 1. 跨团队协作 * 与金融领域研究员深度合作,定义金融行业模型能力的评估体系和模型优化目标。 2. 大规模评测体系构建 * 设计支撑万亿参数、百万亿 token 规模训练的评测框架与流水线; * 结合分布式训练体系(TP / PP / SP / EP / CP),排查评测结果中的系统性噪声与偏差; * 建立模型能力的定量评估框架,支撑架构 / 超参 / 数据 scaling laws 的验证。 3. 评测与训练系统协同 * 深入 Megatron / DeepSpeed / FSDP 等框架底层,理解训练过程对评测结果的潜在…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
Megatron+
https://www.youtube.com/watch?v=hc0u4avAkuM
DeepSpeed+
https://www.youtube.com/watch?v=pDGI668pNg0
FSDP+
https://docs.pytorch.org/tutorials/intermediate/FSDP_tutorial.html
In DistributedDataParallel (DDP) training, each rank owns a model replica and processes a batch of data, finally it uses all-reduce to sync gradients across ranks.
https://www.youtube.com/watch?v=PjEwLgyzuzQ
FSDP provides a comprehensive framework for large model training in PyTorch.
CUDA+
https://developer.nvidia.com/blog/even-easier-introduction-cuda/
This post is a super simple introduction to CUDA, the popular parallel computing platform and programming model from NVIDIA.
https://www.youtube.com/watch?v=86FAWCzIe_4
Lean how to program with Nvidia CUDA and leverage GPUs for high-performance computing and deep learning.
NCCL+
https://developer.nvidia.com/nccl
The NVIDIA Collective Communication Library (NCCL) implements multi-GPU and multi-node communication primitives optimized for NVIDIA GPUs and networking.
还有更多 •••
相关职位
校招2027届蚂蚁星
1. 理论基础扎实: 拥有深厚的深度学习与非凸优化理论功底,具备从第一性原理(FirstPrinciples)对复杂问题进行数学建模和拆解的能力; 2. 架构深度理解: 深入掌握 Transf
更新于 2026-05-14上海|杭州
实习蚂蚁星- Pla
1. 计算机科学、人工智能、数学等相关专业硕士及以上学历; 2. 深入理解 Transformer 架构及 SFT / RLHF / DPO / PPO / GRPO 等核心算法;熟悉 AI Agen
更新于 2026-05-12北京|上海|杭州
校招2027届蚂蚁星
1. 计算机科学、人工智能、数学等相关专业硕士及以上学历; 2. 深入理解 Transformer 架构及 SFT / RLHF / DPO / PPO / GRPO 等核心算法;熟悉 AI Agen
更新于 2026-05-14北京|上海|杭州