
MiniMax大模型算法工程师-Code方向
社招全职算法地点:上海 | 北京状态:招聘
工作描述
我们致力于提升大模型在代码领域的核心能力,让模型更好地理解、生成和分析代码,并通过执行、验证与环境交互解决复杂问题。你将参与代码方向的后训练研发,探索高质量数据、强化学习与反馈驱动的训练方法,持续拓展模型的能力边界。你将负责: 1. 研发面向代码能力的后训练方法,提升模型在代码理解、生成、推理、调试与验证等任务中的表现。 2. 面向复杂代码任务构建训练数据与交互环境,探索高质量数据合成、任务生成与难度控制方法。 3.…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
还有更多 •••
相关职位
社招A182748
1、优秀的代码能力、数据结构和基础算法功底,熟练使用PyTorch、TensorFlow、JAX等任一深度学习框架; 2、熟悉大模型或RL算法和技术,在相关领域有过良好研究记录者优先; 3、在大模型领
更新于 2026-04-21北京
社招A188131
1. 硕士及以上学历,人工智能、计算机科学、电子信息技术等相关专业; 2. 具备以下至少一到两个方向的深入研究或实践经验:大语言模型、多模态大模型、大模型后训练、大模型预训练、强化学习; 3. 深入掌
更新于 2026-06-23深圳
校招研发类
1、计算机科学、软件工程、人工智能、机器学习等相关专业; 2、对未来的大模型技术发展有热情和信心; 3、具备NLP、图生文、语音、Agent相关的算法基础; 4、熟悉Python语言、PyTorch训
更新于 2026-08-18南京|上海|深圳