阿里巴巴集团安全部-行为风控Agent算法专家-杭州
社招全职3年以上技术类-算法地点:杭州状态:招聘
工作描述
任职要求 学历要求 硕士及以上学历,计算机科学、人工智能、自然语言处理、机器学习等相关专业。 经验要求 ● 有实际的Agent设计和落地经验,完整参与过Agent系统的架构设计、开发、上线与迭代。 ● 有行为风控或安全风控业务经验,理解风控业务的核心挑战与技术要求。 技术能力 ● 熟悉Agentic RL算法(如GRPO、PPO、RLHF等),能够通过强化学习推动Agent策略优化。 ● 深入理解大模型技术栈:Transformer架构、Prompt Engineering、RAG、Function Calling、Multi-Agent协作框架。 ● 熟悉主流Agent框架(LangChain、AutoGPT、MetaGPT、CrewAI等),了解其设计思想与适用场景。 ● 扎实的机器学习/深度学习基础,熟悉PyTorch等主流框架。 软素质 ● 具备系统思维,能从业务全局视角设计Agent方案,而非孤立优化单点。 ● 有较强的实验设计与数据分析能力,能通…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
NLP+
https://www.youtube.com/watch?v=fNxaJsNG3-s&list=PLQY2H8rRoyvzDbLUZkbudP-MFQZwNmU4S
Welcome to Zero to Hero for Natural Language Processing using TensorFlow!
https://www.youtube.com/watch?v=R-AG4-qZs1A&list=PLeo1K3hjS3uuvuAXhYjV2lMEShq2UYSwX
Natural Language Processing tutorial for beginners series in Python.
https://www.youtube.com/watch?v=rmVRLeJRkl4&list=PLoROMvodv4rMFqRtEuo6SGjY4XbRIVRd4
The foundations of the effective modern methods for deep learning applied to NLP.
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
系统设计+
https://roadmap.sh/system-design
Everything you need to know about designing large scale systems.
https://www.youtube.com/watch?v=F2FmTdLtb_4
This complete system design tutorial covers scalability, reliability, data handling, and high-level architecture with clear explanations, real-world examples, and practical strategies.
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
社招3年以上
1、硕士研究生及以上学历,计算机、人工智能、软件、信息安全、统计和数学专业优先; 2、3年以上大模型/强化学习相关研发经验,深刻理解RLHF/Agent训练经验; 3、具备业务风控领域(作弊、欺诈、账
更新于 2026-07-14北京
社招1年以上技术类-算法
1、计算机、人工智能、数据科学或相关专业硕士及以上学历 2、3年以上大模型算法经验,精通大模型基本原理与训练方法,有行为风控业务经验优先 3、在大模型、AI或安全顶会(ICLR/NeurIPS/ACL
更新于 2026-06-12北京|杭州
社招3年以上技术类-安全
1.具有攻防经验,熟悉常见漏洞原理,善于漏洞的挖掘、利用 2.熟悉RASP相关技术原理,有 RASP 建设经验者优先 3.具备良好的技术洞察力和创新能力,能够快速识别问题、解决问题 4.具备良好的风险
更新于 2026-06-18北京|杭州