腾讯腾讯云DataBuddy -大模型算法专家
社招全职5年以上腾讯云-大数据技术地点:深圳状态:招聘
工作描述
任职要求 1.计算机、人工智能、数学、软件工程等相关专业硕士及以上学历,具备扎实的机器学习、自然语言处理和大模型基础; 2.熟悉大模型评测、数据构造、SFT、偏好优化、RLHF / RLAIF 等后训练方法,有相关项目实践经验优先; 3.熟悉主流大模型及应用方式,了解模型效果优化、Prompt 评测、上下文工程和工具调用评测方法; 4.具备较强的数据分析和问题归因能力,能够从用户反馈、执行日志和失败案例中提炼高质量训练 / 评测样本; 5.熟练掌握 Python,熟悉 PyTorch / Transformers / vLLM / 数据处理…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
NLP+
https://www.youtube.com/watch?v=fNxaJsNG3-s&list=PLQY2H8rRoyvzDbLUZkbudP-MFQZwNmU4S
Welcome to Zero to Hero for Natural Language Processing using TensorFlow!
https://www.youtube.com/watch?v=R-AG4-qZs1A&list=PLeo1K3hjS3uuvuAXhYjV2lMEShq2UYSwX
Natural Language Processing tutorial for beginners series in Python.
https://www.youtube.com/watch?v=rmVRLeJRkl4&list=PLoROMvodv4rMFqRtEuo6SGjY4XbRIVRd4
The foundations of the effective modern methods for deep learning applied to NLP.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
Prompt+
https://cloud.google.com/vertex-ai/generative-ai/docs/learn/prompts/introduction-prompt-design
A prompt is a natural language request submitted to a language model to receive a response back.
https://learn.microsoft.com/en-us/azure/ai-foundry/openai/concepts/prompt-engineering
These techniques aren't recommended for reasoning models like gpt-5 and o-series models.
https://www.youtube.com/watch?v=LWiMwhDZ9as
Learn and master the fundamentals of Prompt Engineering and LLMs with this 5-HOUR Prompt Engineering Crash Course!
还有更多 •••
相关职位
社招8年以上核心本地商业-基
1. 大学本科及以上学历,计算机、人工智能、数学等相关专业。 2. 扎实的深度学习理论基础,精通PyTorch等主流框架,具备独立复现和改进前沿模型的能力。 3. 在NLP/多模态/强化学习至少一个方
更新于 2026-06-11北京
社招核心本地商业-基
1. 对大模型有技术热情,熟悉GPT/BERT/T5等模型的原理; 2. 熟悉Python,熟练使用TensorFlow/PyTorch/Megatron/Triton等深度学习训练或推理框架,熟悉j
更新于 2026-04-02北京|上海
社招算法开发岗
1.有计算机科学、数学、统计学或相关领域的硕士或博士学位; 2. 熟悉Python与深度学习框架,具有良好的编程能力和扎实的数学理论基础; 3.熟悉掌握大模型相关技术,有实际主导或参与过大模型训练工作
更新于 2025-12-16北京