小红书AI Agent算法实习生
实习兼职大模型地点:北京 | 上海状态:招聘♡ 收藏
工作描述
Qualifications 1、对打造亿级用户使用的 AI Agent 有热情、对 AI 产品有审美 2、具备扎实的机器学习和 NLP 基础,理解大模型技术细节,熟悉 SFT、RL、DPO / PPO / GRPO 等 Post-training 技术栈,熟悉主流的agent harness框架,懂如何做出一流的agentic model。 3、有 AI Agent…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
NLP+
https://www.youtube.com/watch?v=fNxaJsNG3-s&list=PLQY2H8rRoyvzDbLUZkbudP-MFQZwNmU4S
Welcome to Zero to Hero for Natural Language Processing using TensorFlow!
https://www.youtube.com/watch?v=R-AG4-qZs1A&list=PLeo1K3hjS3uuvuAXhYjV2lMEShq2UYSwX
Natural Language Processing tutorial for beginners series in Python.
https://www.youtube.com/watch?v=rmVRLeJRkl4&list=PLoROMvodv4rMFqRtEuo6SGjY4XbRIVRd4
The foundations of the effective modern methods for deep learning applied to NLP.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
GRPO+
https://cameronrwolfe.substack.com/p/grpo
Most early work on RL for LLMs used Proximal Policy Optimization (PPO) as the default RL optimizer, but recent reasoning research relies upon Group Relative Policy Optimization (GRPO).
还有更多 •••
相关职位
实习
1. 计算机相关方向的硕士或博士; 2. 有大模型AI Agent相关研究和项目经历,发表过相关方向的顶会论文优先 3. 具有优秀的解决复杂问题和多人协作沟通的能力,能够独立思考并开展工作,具有强烈的
更新于 2026-07-14北京
实习大模型
1、本科及以上在读,计算机、人工智能、自然语言处理、机器学习、数据科学、数学、统计学等相关专业优先。 2、对大模型、AI Agent、RAG、Tool Use、Prompt Engineering、模
更新于 2026-07-25北京|上海
实习策略算法
1. 硕士及以上学历在读,计算机/AI相关专业,每周出勤 4 天以上,实习期 3 个月以上。 2. 动手能力强,有扎实的编程基础,熟练使用 Vibe Coding 工具,能快速将想法转化为可运行代码。
更新于 2026-08-20北京|上海
