小红书基础模型算法工程师
社招全职3-5年大模型地点:北京 | 上海状态:招聘
工作描述
Qualifications - 有大模型 Pre-training / Mid-training 经验,熟悉训练数据、模型结构、训练稳定性及大规模分布式训练; - 有 Knowledge Distillation / Pre-training Distillation / Model Compression / Pruning 等相关经验; - 有 SFT / RL / RLHF / OPD / Reward Modeling 等 Post-training 经验; - 有 MoE、Attention、Long Context、Speculative Decoding、并行解码或其他模型效率优化方向研究经验; - 熟悉 PyTorch,以及 Megatron-LM、DeepSpeed、FSDP、vLLM、SGLang 等训练或推…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
还有更多 •••
相关职位
社招1-3年大模型
我们希望你具备: 背景: 计算机、视觉、机器人等相关专业硕士/博士;熟悉主流 VLM 架构(如 LLaVA, Qwen-VL, InternVL 等)。 专业深耕: 在 计算机视觉(CV)、多模态学习
更新于 2026-07-07北京|上海
社招1-3年大模型
我们希望你具备: 背景: 计算机、数学等相关专业硕士/博士;深入理解 Transformer 架构及大模型训练全流程。 专业深耕: 在 Search(搜索)、Code(代码生成/工程)、tool-us
更新于 2026-07-07北京|上海|杭州
社招3-5年策略算法
1. 计算机相关专业优先; 2. 具备优秀的编码能力,扎实的数据结构和算法功底,至少熟练java/python/golang/c++其中一种开发语言,熟练掌握tensorflow/pytorch;
更新于 2026-08-10北京|上海|杭州