阿里巴巴数据技术及产品部-大模型数据处理优化专家-杭州/北京
社招全职5年以上技术类-开发地点:北京 | 杭州状态:招聘
工作描述
任职要求 1、精通Python/Java,熟悉PyTorch、vLLM等推理框架,具有处理多模态数据(图像、文本、音频、视频等)的经验,对主流大语言模型推理加速有实践经验者优先; 2、熟悉常用的flink/spark/hadoop等大数据处理框架 3、对Daft、Ray等开源AI数据处理框架熟悉者优先 4、有可观测系统开发或应用经验者优先,包括prometheus、gra…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
Java+
https://www.youtube.com/watch?v=eIrMbAQSU34
Master Java – a must-have language for software development, Android apps, and more! ☕️ This beginner-friendly course takes you from basics to real coding skills.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
vLLM+
https://www.newline.co/@zaoyang/ultimate-guide-to-vllm--aad8b65d
vLLM is a framework designed to make large language models faster, more efficient, and better suited for production environments.
https://www.youtube.com/watch?v=Ju2FrqIrdx0
vLLM is a cutting-edge serving engine designed for large language models (LLMs), offering unparalleled performance and efficiency for AI-driven applications.
Flink+
https://nightlies.apache.org/flink/flink-docs-release-2.0/docs/learn-flink/overview/
This training presents an introduction to Apache Flink that includes just enough to get you started writing scalable streaming ETL, analytics, and event-driven applications, while leaving out a lot of (ultimately important) details.
https://www.youtube.com/watch?v=WajYe9iA2Uk&list=PLa7VYi0yPIH2GTo3vRtX8w9tgNTTyYSux
Today’s businesses are increasingly software-defined, and their business processes are being automated. Whether it’s orders and shipments, or downloads and clicks, business events can always be streamed. Flink can be used to manipulate, process, and react to these streaming events as they occur.
Spark+
[英文] Learning Spark Book
https://pages.databricks.com/rs/094-YMS-629/images/LearningSpark2.0.pdf
This new edition has been updated to reflect Apache Spark’s evolution through Spark 2.x and Spark 3.0, including its expanded ecosystem of built-in and external data sources, machine learning, and streaming technologies with which Spark is tightly integrated.
还有更多 •••
相关职位
社招5年以上技术类-数据
1. 本科及以上学历,计算机相关专业;具备数据处理或模型训练经验。 2. 熟练掌握文本、多模态等非结构化数据处理方法,精通数据清洗、去重、相似度计算、脱敏、特征提取与数据增强等技术。 3. 精通Pyt
更新于 2026-07-29北京|杭州
社招3年以上运营-商业伙伴运
1. 本科及以上学历,3年以上资源运营相关经验,熟悉语音、代码、视频、金融、医疗、理工科等经验者优先;硕博优先; 2. 具备大模型数据标注、评测、数据生产项目经验;有专业技术背景优先,算法背景优先;
更新于 2026-08-05北京|杭州
社招5年以上技术类-数据
1.计算机、数学、统计、人工智能、大数据、机器人等相关专业硕士及以上学历,有数据算法相关实践经验,具备多模态/跨模态数据处理实践经验优先; 2. 有相关大规模分布式系统开发经验,熟悉主流大数据存储处理
更新于 2026-05-22杭州