
MiniMax大模型效率研究员
社招全职基础架构地点:北京 | 上海状态:招聘♡ 收藏
工作描述
我们正在寻找一位能够端到端从模型算法、系统软件与硬件角度思考模型效率的研究员。从模型设计早期出发,需要你和算法及训练、推理团队共同探索模型结构,评估其在真实硬件和业务负载下的效率与成本,推动模型能力、训练效率和推理性能的协同提升,同时关注模型效果。 岗位职责 - 参与模型结构设计:从硬件能力、分布式训练和推理系统的视角,参与 Attention、MoE、条件计算、参数共享、量化及缓存机制等设计,提出可验证的架构方案。 - 建立端到端评估方法:建立计算、访存、通信及存储的性能与成本模型,分析模型结构对训练吞吐、显存占用、收敛表现,以及推理 TTFT、TPOT、吞吐和服务成本的影响。 - 关注效率与模型效果的tradeoff:能够端到端参与到模型效果评估,可以衡量基于效率的变更对效果的影响,从而找到最佳的tradeoff。 - 推动模型与硬件协同设计:结合加速器的计算精度、内存层级、带宽、互联拓扑和执行特性,识别架构瓶颈,为模型规模、层数、…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
缓存+
https://hackernoon.com/the-system-design-cheat-sheet-cache
The cache is a layer that stores a subset of data, typically the most frequently accessed or essential information, in a location quicker to access than its primary storage location.
https://www.youtube.com/watch?v=bP4BeUjNkXc
Caching strategies, Distributed Caching, Eviction Policies, Write-Through Cache and Least Recently Used (LRU) cache are all important terms when it comes to designing an efficient system with a caching layer.
https://www.youtube.com/watch?v=dGAgxozNWFE
内核+
https://www.youtube.com/watch?v=C43VxGZ_ugU
I rummage around the Linux kernel source and try to understand what makes computers do what they do.
https://www.youtube.com/watch?v=HNIg3TXfdX8&list=PLrGN1Qi7t67V-9uXzj4VSQCffntfvn42v
Learn how to develop your very own kernel from scratch in this programming series!
https://www.youtube.com/watch?v=JDfo2Lc7iLU
Denshi goes over a simple explanation of what computer kernels are and how they work, alonside what makes the Linux kernel any special.
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
还有更多 •••
相关职位
社招1-3年J0011
1、本科及以上学历,计算机、数学、统计学、人工智能或相关专业; 2、熟悉任一机器学习分支领域(如统计学习、深度学习、强化学习、组合优化或其他相关前沿技术等); 3、有较强的工程实现能力,熟悉LLM基本
更新于 2026-01-09北京
社招3-5年
1、学历与经验: 计算机、人工智能、数学等相关专业硕士及以上学历,3 年及以上算法或大模型相关工作经验,具备大语言模型在真实业务场景中的落地实践经验者优先。 2、大模型与算法能力: 熟悉主流大语言模
更新于 2026-06-16深圳
社招3-5年
1、学历与经验: 计算机、人工智能、数学等相关专业硕士及以上学历,3 年及以上算法或大模型相关工作经验,具备大语言模型在真实业务场景中的落地实践经验者优先。 2、大模型与算法能力: 熟悉主流大语言模
更新于 2026-06-16深圳