小红书【Hi Lab】大模型AI native算法工程师(垂类)
社招全职大模型地点:北京 | 上海状态:招聘
工作描述
任职要求 1、扎实机器学习与深度学习基础,熟练掌握 PyTorch / JAX / TensorFlow 等任一框架 2、熟悉后训练常用技术(SFT、RLHF / DPO / RLAIF 等)或具备相关项目 / 竞赛 / 论文经验 3、具备 实验设计与问题定位能力,能独立分析大模型在不同数据分布和任务场景下的表现 4、善于沟通和团队协作,乐于在快速迭代中分享想法、推动落地 加分项 - 有深度参与贡献的顶会(ICML / NeurIPS / ICLR / ACL / CVPR 等)论文 - ACM-ICPC、NOI/IOI、Kaggle 等竞赛奖项 - 参…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
JAX+
https://docs.jax.dev/en/latest/notebooks/thinking_in_jax.html
JAX is a library for array-oriented numerical computation, with automatic differentiation and JIT compilation to enable high-performance machine learning research.
TensorFlow+
https://www.youtube.com/watch?v=tpCFfeUEGs8
Ready to learn the fundamentals of TensorFlow and deep learning with Python? Well, you’ve come to the right place.
https://www.youtube.com/watch?v=ZUKz4125WNI
This part continues right where part one left off so get that Google Colab window open and get ready to write plenty more TensorFlow code.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
校招产品运营
1、文理兼修:曾受过心理学/哲学/文学/历史/艺术等方向的专业训练,有顶级的文科素养,同时有较好的逻辑思维能力与批判性思维习惯,能够在捕捉细腻情感与抽象思考间游刃有余; 2、有较好的文字功底:对文字表
更新于 2026-02-10上海
社招1-3年大模型
1. 具备扎实的机器学习基础,能熟练使用至少一种深度学习框架(e.g. PyTorch、Jax、TensorFlow、MindSpore、PaddlePaddle)。 2. 对监督学习、强化学习、表示
更新于 2025-09-26北京|上海|杭州
社招3-5年大模型
具备数据采集/爬虫策略或大规模数据规划经验,熟悉网页、学术、公开语料等主流数据源特性与获取技术。 具备数据价值评估能力,能结合模型训练需求(如稀缺资源、长尾领域)制定数据增强策略。 熟悉数据合规与版权
更新于 2025-10-29北京|上海|广州