滴滴AI算法安全专家(JR2026082400G)
社招全职技术地点:杭州状态:招聘
工作描述
任职要求 1.深入理解大模型原理(Transformer、MoE、SFT、RLHF/DPO),有实际垂域模型(>10B参数)微调与对齐经验,能独立解决训练中的Loss震荡、模型崩溃等问题;对大模型安全对齐、越狱攻防、内容安全、Agent安全中的至少两个方向有系统性实战经验。 2. 对LLM攻击手法有系统性的实战认知。能构造绕过主流安全护栏的Prompt,深刻理解黑灰产在金融场景下的攻击动机与链路,具备以攻击者视角设计防御的思维习惯。 3. 熟练使用Python/Java,熟悉高性能服务框架,能独立将PyTorch模型封装为高并发、低延迟的在线服务,并有实际压测与性能调优经验;熟悉分布式训练框架,了解模型量化、蒸馏等推理优化技术。 …
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
Prompt+
https://cloud.google.com/vertex-ai/generative-ai/docs/learn/prompts/introduction-prompt-design
A prompt is a natural language request submitted to a language model to receive a response back.
https://learn.microsoft.com/en-us/azure/ai-foundry/openai/concepts/prompt-engineering
These techniques aren't recommended for reasoning models like gpt-5 and o-series models.
https://www.youtube.com/watch?v=LWiMwhDZ9as
Learn and master the fundamentals of Prompt Engineering and LLMs with this 5-HOUR Prompt Engineering Crash Course!
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
还有更多 •••
相关职位
社招3年以上技术类-算法
●计算机、电子、数学等相关专业硕士及以上学历; ●3 年以上反爬或安全攻防相关工作经验,有大规模平台反爬实战经验者优先。 ● 有 LLM 在安全或风控场景中的实际落地经历,熟悉大模型的训练、微调、推理
更新于 2026-06-04杭州

社招系统序列
1. 硕士研究生毕业,人工智能方向专业 2. 具备AI算法开发经历和经验,熟悉 CNN、BEV 等人工智能相关内容基础知识,理解主流 ML 框架(如 PyTorch、TensorFlow)的基本工作原
更新于 2026-01-23上海|北京

校招算法类
"1、博士学历; 2、熟练掌握Python,熟悉Linux 环境开发,熟练使用深度学习框架TensorFlow或者PyTorch; 3、熟悉一项或者多项以下技术:LLM预训练、对话管理、Instruc
更新于 2025-11-21北京