快手大模型算法工程师 - 【综效线-效率工程部】
社招全职1-3年J0011地点:北京状态:招聘
工作描述
任职要求 1、本科及以上学历,计算机、数学、统计学、人工智能或相关专业; 2、熟悉任一机器学习分支领域(如统计学习、深度学习、强化学习、组合优化或其他相关前沿技术等); 3、有较强的工程实现能力,熟悉LLM基本原理、大模型微调/RLHF等技术,熟悉C/C++、Python、Java等至少一门主流编程语言; 4…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
社招A188131
1. 硕士及以上学历,人工智能、计算机科学、电子信息技术等相关专业; 2. 具备以下至少一到两个方向的深入研究或实践经验:大语言模型、多模态大模型、大模型后训练、大模型预训练、强化学习; 3. 深入掌
更新于 2026-06-23深圳
校招研发类
1、计算机科学、软件工程、人工智能、机器学习等相关专业; 2、对未来的大模型技术发展有热情和信心; 3、具备NLP、图生文、语音、Agent相关的算法基础; 4、熟悉Python语言、PyTorch训
更新于 2026-08-18南京|上海|深圳
社招算法开发岗
1.扎实的机器学习、NLP、RL基础和出色的创新能力,在ACL/EMNLP/NAACL/NeurIPS/ICML/ICLR等顶级会议上发表论文者优先; 2.在预训练、后训练、强化学习方向有深刻研究;
更新于 2026-06-24北京