快手大模型后端工程师-【可灵AI】
社招全职3-5年J0012地点:北京 | 深圳状态:招聘
任职要求
1、熟练掌握diffusion原理,熟悉transformer结构及其变种,掌握大模型模型特性,有过大模型训练经历,SFT经历者优先; 2、熟练掌握传统模型压缩技术,包括:模型量化,模型稀疏化(如剪枝,token-merge,token-eviction),模型蒸馏,有其中一个相关的研究经历或实践经验; 3、熟练掌握…
登录查看完整任职要求
微信扫码,1秒登录
工作职责
1、负责可灵数字人团队生成端系统,包括技术方案设计、算法对接服务部署、业务方对接工作; 2、负责可灵数据团队内部文本及多模态大模型的推理部署效率优化需求。
包括英文材料
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
缓存+
https://hackernoon.com/the-system-design-cheat-sheet-cache
The cache is a layer that stores a subset of data, typically the most frequently accessed or essential information, in a location quicker to access than its primary storage location.
https://www.youtube.com/watch?v=bP4BeUjNkXc
Caching strategies, Distributed Caching, Eviction Policies, Write-Through Cache and Least Recently Used (LRU) cache are all important terms when it comes to designing an efficient system with a caching layer.
https://www.youtube.com/watch?v=dGAgxozNWFE
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
还有更多 •••
相关职位
社招3-5年J0012
1、配合算法同学,推动深度学习相关算法的落地,打造高吞吐、低延时的推理系统; 2、优化大模型推理服务性能,提升吞吐并控制成本; 3、优化大模型推理服务化框架,提升框架易用性和可调试性。
更新于 2025-12-23北京
社招1年以上A183970B
1、负责字节跳动豆包语音大模型的原子能力后端开发,确保在豆包、剪映、即梦、抖音、番茄、火山引擎等落地; 2、持续推进语音理解、生成创作以及多模态端到端大模型最新技术的工程优化和应用落地; 3、大模型分布式推理的架构和优化,在模型高速迭代的过程中,保证架构可扩展、高可用、合理资源利用率; 4、大模型语音服务稳定性治理,并发优化,多地域运维部署提效; 5、支撑业务,建设输出创新能力提升豆包、剪映、即梦、抖音等产品语音交互和生成创作体验。
更新于 2026-05-07上海
实习后端开发
工作职责: 参与小红书点点离线工具链后台系统的架构设计和开发工作,保障系统稳定,高效,安全; 参与设计开发数据采集、数据标注、数据管理等平台系统,持续优化和迭代工具数据闭环能力; 参与测评平台的快速基建开发,实现测评基建能力的持续提升; 深入发掘和分析业务需求,设计合理的技术方案并实现;
更新于 2026-02-12北京
