小红书AllSpark-AI Agent算法工程师
社招全职3-5年大模型地点:北京 | 上海状态:招聘♡ 收藏
工作描述
Qualifications 1、对打造亿级用户使用的 AI Agent 有热情、对 AI 产品有审美 2、具备扎实的机器学习和 NLP 基础,理解大模型技术细节,熟悉 SFT、RL、DPO / PPO / GRPO 等 Post-training 技术栈,熟悉主流的agent harness框架,懂如何做出一流的agentic model。 3、有 AI Agent…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
NLP+
https://www.youtube.com/watch?v=fNxaJsNG3-s&list=PLQY2H8rRoyvzDbLUZkbudP-MFQZwNmU4S
Welcome to Zero to Hero for Natural Language Processing using TensorFlow!
https://www.youtube.com/watch?v=R-AG4-qZs1A&list=PLeo1K3hjS3uuvuAXhYjV2lMEShq2UYSwX
Natural Language Processing tutorial for beginners series in Python.
https://www.youtube.com/watch?v=rmVRLeJRkl4&list=PLoROMvodv4rMFqRtEuo6SGjY4XbRIVRd4
The foundations of the effective modern methods for deep learning applied to NLP.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
GRPO+
https://cameronrwolfe.substack.com/p/grpo
Most early work on RL for LLMs used Proximal Policy Optimization (PPO) as the default RL optimizer, but recent reasoning research relies upon Group Relative Policy Optimization (GRPO).
还有更多 •••
相关职位
社招3-5年大模型
1、背景: 计算机、电子、数学等相关专业硕士/博士;深入理解大模型训练、推理和数据构建流程; - 2、专业深耕:在预训练(数据配比,模型结构,AI Infra)、SFT(e.g. 数据合成、拒绝采样)
更新于 2026-09-30上海|北京|杭州
社招1-3年大模型
基础素质:计算机、人工智能等相关专业硕士及以上学历,具备扎实的机器学习/深度学习理论基础,精通 Python,熟练使用 PyTorch 等框架。 LLM 经验:深入理解 Transformer 架构,
更新于 2026-09-22上海|北京
社招1-3年大模型
我们希望你具备: 背景: 计算机、视觉、机器人等相关专业硕士/博士;熟悉主流 VLM 架构(如 LLaVA, Qwen-VL, InternVL 等)。 专业深耕: 在 计算机视觉(CV)、多模态学习
更新于 2026-09-30北京|上海