华为AI模型研究员-端侧大模型
校招全职研发类地点:深圳 | 北京 | 上海 | 杭州 | 南京 | 苏州 | 东莞 | 武汉 | 西安 | 成都 | 长沙 | 济南状态:招聘
工作描述
任职要求 1、计算机科学、人工智能、软件工程、电子信息、自动化、统计数学等相关专业,在顶会有高质量学术产出; 2、深入理解 Transformer 架构及主流 LLM 训练技术栈,对大模型底层原理有透彻理解; 3、具备在 MoE、稀疏/线性注意力、KV Cache 优化、投机推理、蒸馏等模型小型化相关研究或实践经验; 4、熟悉语言大模…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
缓存+
https://hackernoon.com/the-system-design-cheat-sheet-cache
The cache is a layer that stores a subset of data, typically the most frequently accessed or essential information, in a location quicker to access than its primary storage location.
https://www.youtube.com/watch?v=bP4BeUjNkXc
Caching strategies, Distributed Caching, Eviction Policies, Write-Through Cache and Least Recently Used (LRU) cache are all important terms when it comes to designing an efficient system with a caching layer.
https://www.youtube.com/watch?v=dGAgxozNWFE
还有更多 •••
相关职位
校招研发类
1、教育背景:计算机科学、人工智能、数据科学、软件工程、统计数学等相关专业; 2、精通Python等编程语言,熟悉深度学习框架,具备软件工程和算法实现能力; 3、具备较强的学习力、自驱力、团队协同和责
更新于 2026-07-23深圳|北京|上海
校招研发类
1、计算机科学、人工智能、数据科学、软件工程、统计数学等相关专业; 2、具备软件工程和算法实现能力,熟悉模型架构、数据工程、深度学习框架和分布式并行软件开发环境; 3、具备较强的学习力、自驱力、团队协
更新于 2026-07-23北京|上海|深圳
校招研发类
1、计算机科学与技术、软件工程、人工智能、控制科学与工程、数据科学与大数据技术、自动化(偏AI控制方向)等相关专业;深入理解主流搜索/推荐模型(如FM、DSSM、DIN、Transformer等),具
更新于 2026-07-23深圳|北京|上海