小红书【REDstar】Dots-Post-Training算法工程师
校招全职地点:北京 | 上海状态:招聘
工作描述
任职要求 1、本科及以上学历,计算机等相关专业方向优先; 2、扎实机器学习与深度学习基础,熟练掌握PyTorch / JAX / TensorFlow等任一框架; 3、熟悉后训练常用技术(SFT、RLHF / DPO / RLAIF 等)或具备相关项目 / 竞赛 / 论文经验; 4、具备实验设计与问题定位能力,能独立分析大模型在不同数据分布和任务场景下的表现; 5、善于沟通和团队协作,乐于在快速迭代中分享想法、推动落地。 【加分项】 1、有深度参与贡献的顶会(ICML / NeurIPS / ICLR / ACL / CVPR 等)论文; 2、ACM-ICPC、NOI/IOI…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
JAX+
https://docs.jax.dev/en/latest/notebooks/thinking_in_jax.html
JAX is a library for array-oriented numerical computation, with automatic differentiation and JIT compilation to enable high-performance machine learning research.
TensorFlow+
https://www.youtube.com/watch?v=tpCFfeUEGs8
Ready to learn the fundamentals of TensorFlow and deep learning with Python? Well, you’ve come to the right place.
https://www.youtube.com/watch?v=ZUKz4125WNI
This part continues right where part one left off so get that Google Colab window open and get ready to write plenty more TensorFlow code.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
校招
1、本科及以上学历,计算机等相关专业方向优先; 2、熟练掌握GPU CUDA编程; 3、熟练掌握Megatron/DeepSpeed/FSDP等框架的研发,熟练掌握模型并行分布式技术; 4、对大语言模
更新于 2026-08-24北京|上海
校招
1、本科及以上学历,计算机等相关专业方向优先; 2、熟练掌握代码的阅读和编写技术能力,善于利用AI编程,在python和pytorch技术栈上有深入实践; 3、在视觉领域有深入的实践和理解,有一定深度
更新于 2026-08-24北京|上海
校招
1、本科及以上学历,计算机、数学、物理等相关专业方向优先; 2、在大规模数据处理、评测、大规模训练、结构设计、优化、对齐、RL等方向的实际经验和差异化想法者优先; 3、在ICLR/NeurIPS/IC
更新于 2026-08-24北京|上海