智谱大模型安全能力研究员(实习)
实习兼职地点:北京状态:招聘♡ 收藏
工作描述
【职位描述】 和现有团队一起研究,提升大模型在漏洞攻防方向的安全能力。 【核心职责】 1. 设计并落地面向漏洞挖掘的大模型训练方案,涵盖预训练语料构建、指令微调以及基于人类/AI 反馈的强化学习环节; 2. 探索提升模型在自动源码审计、逆向工程、动态调试、模糊测试、漏洞利用程序(Exp)构建等安全场景下推理能力的技术路径; 3. 搭建并持续迭代模…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
还有更多 •••
相关职位
社招A41838
岗位要求 - 计算机、网络安全等相关方向硕博学历以上,具备 CTF 参赛经历或漏洞挖掘实战经验; - 熟悉大模型训练全链路(预训练、SFT、RLHF) - 具备扎实的系统编程能力(C/C++/Rust
更新于 2026-04-10北京
实习阿里巴巴2027
1、扎实的机器学习/深度学习基础,熟悉大模型评测方法(LLM-as-a-Judge、红队测试、对抗评测等); 2、深入理解 Claude、OpenAI 等顶尖实验室的模型评测技术体系,具备改进与创新能
更新于 2026-03-23北京|杭州
社招A249359A
1、本科及以上学历,有AI评测产品、模型评测、数据评估、内容质量、AI产品运营等相关工作经验优先; 2、有自动评估相关实践经验,了解LLM Eval的基本框架,理解LLM as a Judge、规则评
更新于 2026-08-04北京