快手大模型创作智能体研究员-【可灵AI专项】
社招全职3-5年J0011地点:北京状态:招聘
工作描述
任职要求 1、数学、计算机、控制科学、软件工程、人工智能等相关学科,硕士研究生及以上学历; 2、熟悉大模型的相关基础知识,具备大语言模型相关训练或推理的基础知识; 3、熟悉LLM的训练或Fine-tuning的方法,例如SFT/RLHF经验,或熟悉强化学习(RL)概念深入了解DPO、PPO相关算法知识; 4、有大模型对齐项目经验,有agent开发、优化经验者优先; 5、扎实的Python或者C++编…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
还有更多 •••
相关职位
实习A248661
1、2027届及以后毕业,博士在读,计算机、人工智能等相关专业优先; 2、具备一个或多个领域的研究、实践经验,包括但不限于以下方向; 1)对多模态理解/Omni-modal模型/LLM的Post-Tr
更新于 2026-04-20上海
实习A240421A
1、2027届及以后毕业,博士在读,计算机、人工智能等相关专业优先; 2、具备一个或多个领域的研究、实践经验,包括但不限于以下方向; 1)对多模态理解/Omni-modal模型/LLM的Post-Tr
更新于 2026-04-20深圳
校招A108105A
1、2027届毕业,获得博士学位,计算机、人工智能等相关专业优先; 2、具备一个或多个领域的研究、实践经验,包括但不限于以下方向; 1)对多模态理解/Omni-modal模型/LLM的Post-Tra
更新于 2026-04-15深圳