logo of amd

AMDAI Framework Engineer

社招全职 Engineering地点:上海状态:招聘

工作描述


任职要求
1. Expertise in Inference Frameworks: Proven, hands-on experience with vLLM or SGLang, including deep understanding of their source code, deployment, configuration, and performance tuning. (Please describe relevant projects in your resume). 2. Mastery of Model Architectures: In-depth understanding and practical experience with inference workflows of mainstream LLMs (e.g., DeepSeek, Qwen), including their tokenizers, model configurations, and architecture definitions. 3. Strong Theoretical Foundation: Solid grasp of the principles behind Transformer, Self-Attention, MoE, KV Cache, and their impact on inference performance. 4. Proven Optimization Experience: Familiarity with end-to-end LLM inference optimization techniques such as PagedAttention, FlashAttention, continuous/dynamic batching, and quantization (INT8/INT4/GPTQ/AWQ), demonstrated with successful case studies. 5. Programming Skills: Proficiency in Python and strong software engineering best practices. Preferred Qualifications (Plus): 1. Low-Level Development Skills: Experience with CUDA C++ programming for writing and debugging high-perf…
登录查看完整工作描述
微信扫码,1秒登录

包括英文材料
大模型+
vLLM+
SGLang+
Transformer+
缓存+
还有更多 •••
相关职位

logo of amd
社招 Enginee

Skilled engineer with strong technical and analytical expertise in C++ development within Linux envi

更新于 2025-10-25上海
logo of amd
社招 Enginee

Skilled engineer with strong technical and analytical expertise in C++ development within Linux envi

更新于 2025-12-09上海
logo of amd
社招 Enginee

Benefits offered are described: AMD benefits at a glance. AMD does not accept unsolicited resumes

更新于 2026-06-24上海
logo of amd
社招 Enginee

Skilled engineer with strong technical and analytical expertise in C++ development within Linux envi

更新于 2025-12-08上海