智谱GLM大模型算法研究员(推理/规划/Agent方向)
社招全职地点:北京状态:招聘
工作描述
【团队介绍】 当前语言模型在通用对话、基础推理和短程任务处理中已经取得了显著进展,但在 Agent 场景下的复杂任务求解 中,模型距离真正稳定、泛化、可扩展地完成真实世界任务仍有明显差距。尤其是在 Deep Research、Code Agent、开放环境决策、多步工具使用、长程任务分解与执行 等问题上,模型常常面临规划深度不足、长期奖励稀疏、探索效率低、任务迁移能力弱等挑战。我们关注如何让模型具备更强的 长程推理、持续探索、自主规划与多阶段决策能力,使其能够在复杂环境中围绕目标进行多轮思考、行动、反馈和修正,逐步完成具有真实价值的 long-horizon tasks。团队重点研究: 面向 Agent 的推理、规划与执行一体化建模 基于强化学习(RL)的复杂任务能力提升 长程任务中的探索、信用分配与奖励建模 面向真实环境的工具使用、代码执行与交互式学习 在…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
智能体+
https://learn.microsoft.com/en-us/shows/ai-agents-for-beginners/
In this 10-lesson course we take you from concept to code while covering the fundamentals of building AI agents.
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
Web+
https://web.dev/learn
Explore our growing collection of courses on key web design and development subjects.
学历+
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
还有更多 •••
相关职位
社招
职位描述 负责强化学习训练框架的研发、优化和维护,根据业务需求持续改进训练框架和策略,提升模型训练效率 分析和定位训练中的性能瓶颈,实施针对性优化措施,提升训练效率和稳定性 跟进业界技术进展,不断同步
更新于 2026-09-08北京
社招1年以上
【岗位职责】 Code Agent 框架开发与迭代:参与公司自研或开源Code Agent框架的设计、开发与迭代,优化其代码理解、工具调用、多轮交互、记忆管理及任务规划等核心能力。 评测体系与基准构建
更新于 2026-08-17北京
实习
【岗位职责】 代表 Z.ai 参与 transformers、vLLM、SGLang、Ollama 等主流开源项目的建设与维护,向开源社区贡献关键代码,推动 GLM 系列开源模型在各大训练/推理框架上
更新于 2026-09-14北京