智谱大模型应用算法工程师
社招全职3年以上地点:北京状态:招聘
工作描述
一、工作职责: 负责 LLM微调(SFT/LoRA 等)、对齐优化(DPO/RLHF 等)全流程,优化模型性能、效率及鲁棒性,攻克训练与推理核心技术难点; 独立设计并落地大模型在对话系统、内容生成、agent应用、端侧多模态大模型等业务场景的算法方案,优化准确率、生成质量等核心指标; 跟踪 NLP/LLM 领…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
NLP+
https://www.youtube.com/watch?v=fNxaJsNG3-s&list=PLQY2H8rRoyvzDbLUZkbudP-MFQZwNmU4S
Welcome to Zero to Hero for Natural Language Processing using TensorFlow!
https://www.youtube.com/watch?v=R-AG4-qZs1A&list=PLeo1K3hjS3uuvuAXhYjV2lMEShq2UYSwX
Natural Language Processing tutorial for beginners series in Python.
https://www.youtube.com/watch?v=rmVRLeJRkl4&list=PLoROMvodv4rMFqRtEuo6SGjY4XbRIVRd4
The foundations of the effective modern methods for deep learning applied to NLP.
还有更多 •••
相关职位
实习高德研究型实习生
1、计算机相关方向本科及以上学历。 2、良好的编程能力,熟悉主流深度学习工具PyTorch/TensorFlow等。 3、熟悉常见的机器学习和NLP算法,了解当前LLM热点和前沿技术,有过LLM相关的
更新于 2025-12-05北京
实习核心本地商业-业
1. 计算机、数学、统计、信息科学等相关专业的本科或以上学历,扎实的数学、统计学基础; 2. 熟练掌握机器学习和深度学习算法,熟练运用 SQL、Python等工具; 3. 突出的分析问题和解决问题能力
更新于 2026-01-22北京
社招3-5年J0011
1、本硕博学历均可;计算机、人工智能、数学相关专业; 2、有较强的工程实现能力,熟悉LLM及MLLM基本原理、大模型微调/RLHF等技术,熟悉C/C++、Python、Java等至少一门主流编程语言;
更新于 2025-12-22北京