字节跳动豆包大模型算法工程师(AIGC Agent方向) - Data AML
校招全职A132162A地点:杭州 | 北京 | 上海 | 深圳状态:招聘
工作描述
任职要求 1、2027届获得本科及以上学历,专业不限; 2、熟悉LLM/VLM/Agent相关技术,具备大模型训练、后训练或Agent优化相关经验;具备Agent RL能力优化;具备AIGC、多模态生成、智能体等方向实践经验,对AI创作场景有长期兴趣和深入理解;具备Agent Post-training、Agent Evaluation、Agent Loop优化经验;具备Multi-Agent、Context Management、Memory、自我进化Agent等方向的研究或工程经验; 3、具备模型迭代经验,深入参与以下任一方向的研究或工程实践: 1)RFT/RLHF/GRPO/OPD等模型强化与对齐优化技术; 2)SFT训练体系建设与高质量数据构建; 3)多模态模型后训练与能力提升; 4、了解图像/视频生成模型技术,包括Diffusi…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
AIGC+
https://ui.adsabs.harvard.edu/abs/2023arXiv230406632W/abstract
To address the challenges of digital intelligence in the digital economy, artificial intelligence-generated content (AIGC) has emerged.
智能体+
https://learn.microsoft.com/en-us/shows/ai-agents-for-beginners/
In this 10-lesson course we take you from concept to code while covering the fundamentals of building AI agents.
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
NeurIPS+
https://neurips.cc/
ICML+
https://icml.cc/
ICLR+
https://iclr.cc/
还有更多 •••
相关职位
校招A197621A
1、2027届获得本科及以上学历; 2、对Agent Harness有技术热情,对大模型和Agent有较深入的理解,有Agent Harness的优化经验;有使用AI Agent工具进行软件开发的经验
更新于 2026-08-05北京|上海|杭州
校招A25547
1、2027届获得本科及以上学历; 2、对Agent Harness有技术热情,对大模型和Agent有较深入的理解,有Agent Harness的优化经验;有使用AI Agent工具进行软件开发的经验
更新于 2026-07-31上海|北京|深圳
社招A150641A
1、本科及以上学历,计算机、通信、人工智能等相关专业优先; 2、具备扎实的编程基础(Python/Java等),能够独立完成代码开发与调试,具备全栈开发能力; 3、熟悉大模型技术原理,掌握大模型效果与
更新于 2025-12-09杭州