阿里巴巴数据技术及产品部-大模型安全对抗专家-杭州
社招全职2年以上技术类-安全地点:杭州状态:招聘
工作描述
任职要求 1.计算机、网络安全、人工智能等相关专业硕士及以上学历,具备扎实的大模型安全、AI对抗样本、密码学或传统Web安全背景,拥有成熟的模型红队测试或对齐算法落地经验。 2.熟练掌握PyTorch等深度学习框架,在Transformer架构、大模型微调(PEFT/SFT)、RLHF/DPO对齐、强化学习等至少一个领域有深入的算法落地经验。 3.熟悉大模型主流安全漏洞与绕过机制,有基于LLM-as-a-Judge的安全评估框架研发经历,或熟悉多模态对抗样本生成、黑盒/白盒对抗攻击算法者优先。 4.具备优秀的数学建模与安全策略设计能力,熟悉海量攻击等特征工程,对分布式训练、大规模红队自动化工程化实现及推理加速优化有深刻理解。 5.具备强烈的技术创新意识、极客精神与自驱…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
Web+
https://web.dev/learn
Explore our growing collection of courses on key web design and development subjects.
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
社招3年以上技术类-数据
1. 计算机、信息安全、物理、数学相关专业硕士及以上学历。 2. 具备优秀的工程能力,熟悉至少一种主流开发语言(Go、Java、Python 等)。 3. 扎实的密码学基础,熟悉对称/非对称加密、密钥
更新于 2026-09-02杭州
社招3年以上技术类-数据
● 本科及以上学历,网络空间安全、信息安全、计算机科学相关专业;硕士优先。 ● 3 年以上网络安全从业经验,具备以下至少一项深度背景:渗透测试、漏洞研究、安全运营(SOC)、应急响应、安全合规/审计、
更新于 2026-08-06北京|杭州
社招2年以上技术类-算法
1、硕士及以上学历,计算机科学与技术、数学等相关专业优先; 2、有2年以上大模型相关算法研发经验,熟悉常见的深度学习框架(如Tensorflow、Pytorch),有Agentic大模型开发经验优先;
更新于 2026-09-07北京|杭州