快手AI应用算法工程师(AIGC方向)-【商业化】
社招全职1-3年J0011地点:北京状态:招聘
工作描述
任职要求 1、计算机、人工智能、数学相关专业; 2、熟悉大模型的相关基础知识,具备大语言模型相关训练或推理的基础知识;具备通过demo快速验证想法的能力; 3、熟悉LLM的训练或Fine-tuning的方法,例如SFT/RLHF经验,或熟悉强化学习(RL)概念深入了解DPO、PPO相关算法知识; 4、扎实的Python或者C++编程功底,了解PyTorch,Tensorflow,Deepspeed,Megatron,vLLM等大模型训练…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
还有更多 •••
相关职位
社招1-3年J0011
1、计算机科学、数据科学、人工智能、数学等相关专业,具备较好的数据分析和统计学基础,有根据数据表现驱动业务优化的经验; 2、在多模态生成、多模态理解等相关领域有深入的理解,有实际项目经验; 3、优秀的
更新于 2026-07-07北京
实习阿里巴巴2027
1.基础条件 ● 计算机、数学、统计学等相关专业硕士/博士优先,优秀本科生不受限制。 ● 有顶会论文(ACL/EMNLP/ICLR/NeurIPS/ICML等)/高影响项目/开源贡献者加分。 2.专
更新于 2026-03-19北京|广州|杭州
实习阿里巴巴日常实习
1.基础条件 ● 计算机、数学、统计学等相关专业硕士/博士优先,优秀本科生不受限制。 ● 有顶会论文(ACL/EMNLP/ICLR/NeurIPS/ICML等)/高影响项目/开源贡献者加分。 2.专业
更新于 2026-04-03杭州
