腾讯微信读书/输入法/秒剪-大模型评测算法工程师-Agent方向
社招全职1年以上微信读书技术地点:北京状态:招聘
工作描述
任职要求 1.计算机科学、人工智能、数学、统计学等相关专业硕士及以上学历; 2.精通 Python,熟悉 PyTorch/HuggingFace 生态。深入理解 Transformer 架构及大模型训练流程(预训练、SFT、RLHF/DPO); 3.熟悉主流评测框架(如 OpenCompass、lm-evaluation-harness、HF Evaluate 等)及常用指标(BLEU、ROUGE、Pa…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
社招1年以上微信秒剪技术
1.具备NLP或多模态领域算法基础,能深刻理解大模型工作原理; 2.具备优秀的数据分析与处理能力,对数据敏感,能针对不同模型特点制定数据策略; 3.熟练使用 Python 及 Pandas、NumPy
更新于 2026-06-29北京
社招3年以上微信秒剪技术
1.计算机科学、机器学习、人工智能等相关专业,硕士及以上学历; 2.在大模型领域有多个研究成果或落地成果; 3.熟悉各种深度学习框架,如 TensorFlow 或 PyTorch;了解分布式训练框
更新于 2026-06-29北京
社招1年以上微信读书技术
1.计算机相关专业本科及以上学历,一年以上客户端应用开发经验; 2.熟悉 Objective-C / Swift 或 Java / Kotlin ,熟悉 iOS 或 Android 开发,有扎实的算
更新于 2026-07-16广州