阿里巴巴面向真实安全攻防场景的自主智能体(Security AI Agent)体系构建与核心能力研究-阿里星
实习兼职阿里巴巴2027届实习生地点:杭州状态:招聘
工作描述
任职要求 1、在安全与 AI 交叉领域有研究或实践经历(如基于 LLM 的漏洞检测、AI 辅助逆向分析、自动化渗透测试等); 2、具备安全工具使用或安全系统开发经验,理解真实安全运营场景的工作流与痛点; 3、具备独立研究能力,能够阅读并跟进相关领域前沿论文,将研究思路转化为可落地的技术方案; 4、理解主流大语言模型的原理与架构(Transformer、预训练-微调范式、RLHF 等); 5、有大模型微调、推理优化或 Agent 系统(如 ReAct、Tool Use、Planning)的实践经验,有强化学习工程落地经验,尤其是在交互式环境中训练 Agent 的实践; 6、有 Agent 框架开发经验(如 LangChain、AutoGPT、OpenAI Function Calli…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
React+
[英文] Quick Start - React
https://react.dev/learn
This page will give you an introduction to 80% of the React concepts that you will use on a daily basis.
https://www.youtube.com/watch?v=SqcY0GlETPk
Master React 18 with TypeScript! ⚛️ Build amazing front-end apps with this beginner-friendly tutorial.
https://www.youtube.com/watch?v=x4rFhThSX04
Learn modern React basics in the most interactive, hands-on way possible in the full course for beginners.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
还有更多 •••
相关职位
实习核心本地商业-基
1、硕士及以上学历,计算机、人工智能、数学、自然语言处理等相关专业,博士优先; 2、在大模型领域有研究基础,或参与过有影响力的开源项目,在ICLR/NeurIPS/ICML/ACL等顶会发表论文者优先
更新于 2026-04-03北京|上海
实习核心本地商业-基
1、硕士及以上学历,计算机、人工智能、数学、自然语言处理等相关专业,博士优先; 2、在大模型领域有研究基础,或参与过有影响力的开源项目,在ICLR/NeurIPS/ICML/ACL等顶会发表论文者优先
更新于 2026-04-03北京|上海