智谱Trading Agent 后训练算法工程师
社招全职地点:北京状态:招聘
工作描述
岗位职责 构建面向量化研究、因子挖掘和策略开发的 Agent 环境、Harness、任务与评测体系; 探索 AutoResearch 范式,使 Agent 能够自主提出研究假设、编写代码、运行实验、完成回测、分析结果并持续迭代; 设计和构造 Trading Agent 后训练数据,包括轨迹数据、合成数据、偏好数据和失败案例; 研究 Reward 设计、强化学习、拒绝采样、蒸馏等后训练方法,提升模型在长程研究任务中的能力; 分析 Agent …
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
智能体+
https://learn.microsoft.com/en-us/shows/ai-agents-for-beginners/
In this 10-lesson course we take you from concept to code while covering the fundamentals of building AI agents.
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
学历+
还有更多 •••
相关职位
校招
岗位职责 构建面向量化研究、因子挖掘和策略开发的 Agent 环境、Harness、任务与评测体系; 探索 AutoResearch 范式,使 Agent 能够自主提出研究假设、编写代码、运行实验、完
更新于 2026-08-14北京
实习
你可能参与: 构建量化研究、策略开发与交易相关的 Agent 专项环境和评测体系; 探索 AutoResearch 范式,让 Agent 自主完成数据分析、提出假设、挖掘因子与策略、运行回测、分析结果
更新于 2026-08-14北京
社招金融类
Key responsibilities:Monitor and optimize order-book liquidity in real timeOn-board, oversee and inc
更新于 2026-04-02香港