蚂蚁金服蚂蚁集团-Agent 评测研发工程师-健康事业群
社招全职3年以上技术类-算法地点:上海 | 杭州状态:招聘
工作描述
任职要求 1. 计算机/数学/统计/AI 相关专业本科及以上,硕士/博士优先。 2. 扎实的算法与工程能力:熟练 Python,能独立完成评测框架开发、数据处理与分析。 3. 深刻理解 LLM/Transformer 及后训练范式:SFT、RLHF/DPO/Reward Modeling 等;理解 Agent/RAG 的典型失败模式与评测要点。 4. 有大模型评测/后训练经验:大模型评测、后训练、A/B 实验、统计推断等方向有实战沉淀。 5. 熟悉或有经验构建/使用主流评测集与框架者优先:C-Eval、CMMLU、HumanEval、AGIEval,或 HELM、lm-evaluation-harness 等。 6. 有医疗/合规/安全相关经验加分;有大厂大模型算法/评测/平台化经验加分;发表过顶会论文(ICLR/NeurIP…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
社招A219847
1、本科及以上学历,计算机、软件工程、人工智能等相关专业,硕士学位优先; 2、熟练掌握Python,同时熟悉Go/Nodejs/TypeScript中至少一门,具备良好的系统设计能力; 3、熟悉主流大
更新于 2026-07-30北京|深圳
社招2年以上云智能集团
1. 计算机、人工智能、软件工程或相关专业,本科及以上学历,3年以上测试开发或质量保障相关工作经验,有大模型/AI Agent项目经验者优先。 2. 精通Python/Java至少一种编程语言,熟悉主
更新于 2026-08-17杭州
社招3-5年J0005
1、本科及以上学历,熟悉大模型、Agent 的基本原理; 2、有 Agent 评测集设计经验,有良好的沟通能力和团队协作能力; 3、具备良好的抽象能力,能够将 Agent 能力转化为可评测的任务与指标
更新于 2026-04-07北京