字节跳动代码大模型算法专家/Leader-AI应用与创新
社招全职5年以上A159895地点:杭州状态:招聘
工作描述
任职要求 1、计算机、软件工程相关专业本科及以上学历,5年以上算法研究与开发经验; 2、具备扎实的算法基础,包括但不限于LLM、CodeLLM、强化学习等领域的认知和实践经验; 3、精通LLM微调、预训练、推理优化技术,熟悉Agent、RAG等应用范式,有CodeLLM、CodeAgent落地经验者优先; 4、作风正派、责任心强、协作能力好,对工作充满热情,能够深刻影响其他人;乐于亲自下一线,带领团队深入业务共创拿结果。 加分项: 1、有过CodeLLM基座研…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
还有更多 •••
相关职位
实习
【职位描述】 1. 代码预训练数据的清洗和构建,代码能力评估方法的构建,科学、全面地提升 GLM 预训练模型的代码能力; 2. 在后训练阶段探究通过 SFT / RL 等方法提升模型通用代码指令、前端
更新于 2026-05-06北京
社招1年以上技术类-算法
1. 计算机科学、人工智能、机器学习等领域的博士/硕士毕业生。 2. 对上述前沿问题有持续热情,具备独立思考能力和系统性研究思维,敢于挑战现有范式,能够独立应用技术解决复杂问题。 3. 在上述方向有过
更新于 2026-04-02北京|杭州|上海

社招1年以上技术类-算法
1. 计算机科学、人工智能、机器学习等领域的博士/硕士毕业生。 2. 对上述前沿问题有持续热情,具备独立思考能力和系统性研究思维,敢于挑战现有范式,能够独立应用技术解决复杂问题。 3. 在上述方向有过
更新于 2026-04-02北京|杭州|上海