
美图高性能计算工程师 (北京)
社招全职研发类地点:北京状态:招聘♡ 收藏
工作描述
你将深入参与的事: 美图的AI影像正在用Diffusion、DiT、VLM等大模型重写图像与视频创作体验。我们正把模型推进到大规模的线上推理场景,而你,将是那个让每一帧都跑得又快又省的人。 你将具体负责: · 主导图像/视频生成大模型的全链路推理优化——覆盖Diffusion、ViT、DiT、VLM,从端到端延迟、GPU吞吐到算力成本,把每一条推论的性价比压榨到极致; · 深入模型内部定位性能瓶颈,用图优化、算子融合、显存复用、动态shape适配、预处理流水线加速等手段,持续拉低延迟水位; · 落地PTQ/QAT量化(INT4/INT8/FP8)、蒸馏、剪枝、LoRA加速等压缩方案,在画质与推理效率之间做出漂亮的权衡; · 基于TensorRT、vLLM、SGLang、Diffusers等推理后端…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
TensorRT+
https://docs.nvidia.com/deeplearning/tensorrt/latest/getting-started/quick-start-guide.html
This TensorRT Quick Start Guide is a starting point for developers who want to try out the TensorRT SDK; specifically, it demonstrates how to quickly construct an application to run inference on a TensorRT engine.
vLLM+
https://www.newline.co/@zaoyang/ultimate-guide-to-vllm--aad8b65d
vLLM is a framework designed to make large language models faster, more efficient, and better suited for production environments.
https://www.youtube.com/watch?v=Ju2FrqIrdx0
vLLM is a cutting-edge serving engine designed for large language models (LLMs), offering unparalleled performance and efficiency for AI-driven applications.
SGLang+
[英文] Install SGLang
https://docs.sglang.ai/get_started/install.html
SGLang is a fast serving framework for large language models and vision language models.
https://github.com/sgl-project/sgl-learning-materials
缓存+
https://hackernoon.com/the-system-design-cheat-sheet-cache
The cache is a layer that stores a subset of data, typically the most frequently accessed or essential information, in a location quicker to access than its primary storage location.
https://www.youtube.com/watch?v=bP4BeUjNkXc
Caching strategies, Distributed Caching, Eviction Policies, Write-Through Cache and Least Recently Used (LRU) cache are all important terms when it comes to designing an efficient system with a caching layer.
https://www.youtube.com/watch?v=dGAgxozNWFE
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
学历+
C+++
https://www.learncpp.com/
LearnCpp.com is a free website devoted to teaching you how to program in modern C++.
https://www.youtube.com/watch?v=ZzaPdXTrSb8
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
还有更多 •••
相关职位

社招5年以上技术-基础平台
1. 计算机、电子工程或相关专业本科及以上学历,对计算机体系结构有深刻理解。 2. 拥有深厚的GPU/NPU/XPU高性能计算优化经验,精通至少一种异构计算平台及编程模型(如CUDA, ROCm, O
更新于 2026-04-08北京|杭州
社招5年以上技术-基础平台
1. 计算机、电子工程或相关专业本科及以上学历,对计算机体系结构有深刻理解。 2. 拥有深厚的GPU/NPU/XPU高性能计算优化经验,精通至少一种异构计算平台及编程模型(如CUDA, ROCm, O
更新于 2026-07-30北京|杭州
社招2年以上
【我们期待这样的你】 基础扎实:计算机相关专业背景,拥有深厚的 C++ 功底,精通数据结构、算法及操作系统原理(内存管理、多线程、锁机制)。 专项特长(满足以下任一方向即可,不必全能): 推理优化方向
更新于 2026-09-20北京