智谱GLM-强化学习训练框架工程师(RL)
社招全职地点:北京状态:招聘
包括英文材料
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
NLP+
https://www.youtube.com/watch?v=fNxaJsNG3-s&list=PLQY2H8rRoyvzDbLUZkbudP-MFQZwNmU4S
Welcome to Zero to Hero for Natural Language Processing using TensorFlow!
https://www.youtube.com/watch?v=R-AG4-qZs1A&list=PLeo1K3hjS3uuvuAXhYjV2lMEShq2UYSwX
Natural Language Processing tutorial for beginners series in Python.
https://www.youtube.com/watch?v=rmVRLeJRkl4&list=PLoROMvodv4rMFqRtEuo6SGjY4XbRIVRd4
The foundations of the effective modern methods for deep learning applied to NLP.
OpenCV+
https://learnopencv.com/getting-started-with-opencv/
At LearnOpenCV we are on a mission to educate the global workforce in computer vision and AI.
https://opencv.org/university/free-opencv-course/
This free OpenCV course will teach you how to manipulate images and videos, and detect objects and faces, among other exciting topics in just about 3 hours.
还有更多 •••
相关职位
社招1年以上
【岗位职责】 Code Agent 框架开发与迭代:参与公司自研或开源Code Agent框架的设计、开发与迭代,优化其代码理解、工具调用、多轮交互、记忆管理及任务规划等核心能力。 评测体系与基准构建
更新于 2026-08-17北京
实习
【岗位职责】 代表 Z.ai 参与 transformers、vLLM、SGLang、Ollama 等主流开源项目的建设与维护,向开源社区贡献关键代码,推动 GLM 系列开源模型在各大训练/推理框架上
更新于 2026-09-14北京
校招
【岗位职责】 1. Code Agent 框架开发与迭代:参与公司自研或开源Code Agent框架的设计、开发与迭代,优化其代码理解、工具调用、多轮交互、记忆管理及任务规划等核心能力。 2. 评测体
更新于 2026-09-14北京