蚂蚁金服大模型安全与对齐
校招全职蚂蚁集团2027届应届生招聘地点:北京 | 上海 | 杭州状态:招聘
工作描述
任职要求 1. 具备扎实的机器学习与深度学习理论基础;在大模型安全对齐领域有深入研究,熟悉RLHF、SFT、可控生成等核心技术,具备相关算法研发经验;在Agent安全相关领域有深入研究; 2. 精通PyTorch/Megatron/VeRL等深度学习框架,有大模型训练或安全优化实战经验;精通计算机原理,大模型算法、大模型应用系统架构和相关优化经验; 3. 具备优秀的算法实现、系统优化、及工程落地能力; 4. 对AI安全技术发展趋势有深刻理解…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
还有更多 •••
相关职位
实习阿里巴巴2027
1.计算机、数学、人工智能、网络安全等相关专业硬士及以上学历,3年以上算法研发或安全研发经验,有智能体安全、系统安全、大模型安全或偏好对齐领域背景者优先; 2.精通Python/C++/Java/Go
更新于 2026-05-11北京|杭州
实习阿里巴巴日常实习
1、硕士及以上学历在读,计算机、人工智能、网络安全、数学等相关专业优先。 2、具备扎实的深度学习基础,熟悉 Qwen、DeepSeek、GLM、Gemma 等主流大模型架构及训练/推理机制。 3、熟悉
更新于 2026-06-30杭州
实习阿里巴巴日常实习
1、硕士及以上学历在读,计算机、人工智能、网络安全、数学等相关专业优先。 2、具备扎实的深度学习基础,熟悉主流多模态大模型(如Qwen-VL, LLaVA, InternVL等)及视觉生成模型(如St
更新于 2026-06-30杭州