小红书策略算法实习生
实习兼职策略算法地点:北京 | 上海状态:招聘
任职要求
1、本科及以上学历,计算机、软件工程、数学等相关专业; 2、必备优秀的编码能力与算法基础,熟练掌握至少一门主流编程语言(如Python/Go/C++/Java),数据结构与算法功底扎实; 3、强烈的求知欲与卓越的思维潜力,对强化学习、大模型等技术怀有浓厚兴趣,并愿意深入探索,并且具备清晰的逻辑思维,能够对复杂问题进行有…
登录查看完整任职要求
微信扫码,1秒登录
工作职责
1、参与核心策略设计与实现:深入小红书音视频、直播、图片等内容的分发与体验优化全链路,参与转码、下发、消费等核心策略的设计、编码与迭代,并且可以设计清晰、可扩展的技术方案,并通过高质量的代码实现它; 2、有数据挖掘和数据分析能力,并且可以在真实的业务场景中,学习并运用AB实验、因果推断等科学方法评估策略效果。同时探索强化学习、大模型等前沿技术在用户体验优化领域的应用可能。
包括英文材料
学历+
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
Go+
https://www.youtube.com/watch?v=8uiZC0l4Ajw
学习Golang的完整教程!从开始到结束不到一个小时,包括如何在Go中构建API的完整演示。没有多余的内容,只有你需要知道的知识。
C+++
https://www.learncpp.com/
LearnCpp.com is a free website devoted to teaching you how to program in modern C++.
https://www.youtube.com/watch?v=ZzaPdXTrSb8
Java+
https://www.youtube.com/watch?v=eIrMbAQSU34
Master Java – a must-have language for software development, Android apps, and more! ☕️ This beginner-friendly course takes you from basics to real coding skills.
数据结构+
https://www.youtube.com/watch?v=8hly31xKli0
In this course you will learn about algorithms and data structures, two of the fundamental topics in computer science.
https://www.youtube.com/watch?v=B31LgI4Y4DQ
Learn about data structures in this comprehensive course. We will be implementing these data structures in C or C++.
https://www.youtube.com/watch?v=CBYHwZcbD-s
Data Structures and Algorithms full course tutorial java
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
还有更多 •••
相关职位
实习技术类
业务及团队介绍:POI(Point of Interest),即“兴趣点”,是地理信息系统中的重要概念,表示物理世界中的一处地方,可以是一家美食店、一个小区、一栋大楼等。团队通过人工采集、司乘反馈、用户上报、图像等多模态数据,对POI数据进行更新。随着大模型能力的提升,大模型可以更快、更准确的发现POI相关的变化,并辅助人工/自动化 的对数据进行更新。借助大模型能力和完善的数据生态,实时发现物理时间的变化,帮助用户更快更准的找到目的地。
更新于 2025-09-22北京
实习MEG
1、负责大模型智能体(Agent)核心策略研发、迭代与落地 2、参与 Agent 模型训练、评估及 DPO/GRPO 等后训练技术落地 3、参与搭建具备创作感知、文案创作、多能力调度的创作智能体(Creative Agent),支持复杂内容生产流的逻辑编排 4、跟踪并复现 Agent 领域前沿技术(如:Long-context RAG、多模态理解与对齐、DeepResearch 调研框架、多 Agent 协同体系),并在场景中验证其业务价值
更新于 2026-03-25北京
实习阿里巴巴研究型实
1. 负责广告&自然结果在Generator-Evaluator架构下混排模型&机制迭代创新; 2. 负责研究强化学习、生成式模型在广告生成式拍卖中的应用; 3. 负责复杂外部性环境下的广告Listwise价值的预估任务; 4. 负责广告多阶段的分配模型的设计和效率优化; 5. 支持研究和推动Autobidding模式下的新机制设计的应用与创新。
更新于 2026-03-20北京