阿里巴巴企业智能事业部-AI数据研发专家-数据平台
社招全职2年以上技术类-数据地点:杭州状态:招聘
工作描述
任职要求 1. 计算机科学、人工智能、数据科学、信息管理或相关专业硕士及以上学历;具备至少AI算法应用或AI工程化能力一项。 2. 5年以上大数据/AI相关研发经验,其中至少2年聚焦于企业级数据平台、AI应用或大模型数据工程化项目; 3. 具备扎实的数据工程能力,精通SQL,熟练掌握Python等至少一门开发语言,熟悉主流大数据技术栈(如Spark、Flink、Kafka、Hudi/Iceberg)及云数据仓库(如MaxCompute、Hadoop等); 4. 深入理解大模型技术原理,包括预训练、微调、推理等,熟悉主流LLM框架和应用范式(如RAG、Agent等),有实际支持NLP、多模态或AI项目的经验; 5. 对企业经营管理业务有敏锐洞察力,熟悉HR、财务、法务或协同办公等至少一个领域的企业数据逻辑与业务痛点; 6. 具备优秀的系统设计能力与工程落地经验,能独立完成从需求分析到架构设计、开发交付的全过程; 7. 具备扎实的AI工程能力,熟悉大模型应用到生产部署的全流程,能够独立完成Prompt工程、RAG链路设计、Agent任务编排、模型服务封装及线上稳定性治理…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
数据科学+
https://roadmap.sh/ai-data-scientist
Step by step roadmap guide to becoming an AI and Data Scientist
学历+
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
大数据+
https://www.youtube.com/watch?v=bAyrObl7TYE
https://www.youtube.com/watch?v=H4bf_uuMC-g
With all this talk of Big Data, we got Rebecca Tickle to explain just what makes data into Big Data.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SQL+
https://liaoxuefeng.com/books/sql/introduction/index.html
什么是SQL?简单地说,SQL就是访问和处理关系数据库的计算机标准语言。
https://sqlbolt.com/
Learn SQL with simple, interactive exercises.
https://www.youtube.com/watch?v=p3qvj9hO_Bo
In this video we will cover everything you need to know about SQL in only 60 minutes.
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
Spark+
[英文] Learning Spark Book
https://pages.databricks.com/rs/094-YMS-629/images/LearningSpark2.0.pdf
This new edition has been updated to reflect Apache Spark’s evolution through Spark 2.x and Spark 3.0, including its expanded ecosystem of built-in and external data sources, machine learning, and streaming technologies with which Spark is tightly integrated.
Flink+
https://nightlies.apache.org/flink/flink-docs-release-2.0/docs/learn-flink/overview/
This training presents an introduction to Apache Flink that includes just enough to get you started writing scalable streaming ETL, analytics, and event-driven applications, while leaving out a lot of (ultimately important) details.
https://www.youtube.com/watch?v=WajYe9iA2Uk&list=PLa7VYi0yPIH2GTo3vRtX8w9tgNTTyYSux
Today’s businesses are increasingly software-defined, and their business processes are being automated. Whether it’s orders and shipments, or downloads and clicks, business events can always be streamed. Flink can be used to manipulate, process, and react to these streaming events as they occur.
Kafka+
https://developer.confluent.io/what-is-apache-kafka/
https://www.youtube.com/watch?v=CU44hKLMg7k
https://www.youtube.com/watch?v=j4bqyAMMb7o&list=PLa7VYi0yPIH0KbnJQcMv5N9iW8HkZHztH
In this Apache Kafka fundamentals course, we introduce you to the basic Apache Kafka elements and APIs, as well as the broader Kafka ecosystem.
Hudi+
[英文] Spark Quick Start
https://hudi.apache.org/docs/quick-start-guide
we will walk through code snippets that allows you to insert, update, delete and query a Hudi table.
https://www.oreilly.com/library/view/apache-hudi-the/9781098173821/
Overcome challenges in building transactional guarantees on rapidly changing data by using Apache Hudi.
https://www.youtube.com/watch?v=pyK18sDYnS0
In this video, I'll introduce you to one of the most popular Data Lake solutions out there, Apache Hudi!
Iceberg+
https://iceberg.apache.org/spark-quickstart/
This guide will get you up and running with Apache Iceberg™ using Apache Spark™, including sample code to highlight some powerful features.
https://www.baeldung.com/apache-iceberg-intro
This tutorial will discuss Apache Iceberg, a popular open table format in today’s big data landscape.
https://www.youtube.com/watch?v=TsmhRZElPvM
You’ve probably heard about Apache Iceberg™—after all, it’s been getting a lot of buzz.
还有更多 •••
相关职位
社招3年以上技术类-开发
1、本科及以上学历,理科/工科类专业,3年以上Java/python开发经验,具备优秀的逻辑分析能力和良好的沟通协作能力; 2、精通Java/python编程,熟悉常用的设计模式;熟练使用Java/p
更新于 2026-07-31杭州
社招2年以上产品类-平台型
1、硕士及以上学历,计算机科学、人工智能、人力资源管理、管理信息系统等相关专业优先,1年以上产品经理相关工作经验。 2、深刻理解LLM与AI Agent架构原理(如ReAct、Tool Use、Pla
更新于 2026-08-12杭州
社招3年以上技术类-开发
1、本科及以上学历,理科/工科类专业,2年以上Javascript/Java/python开发经验,具备优秀的逻辑分析能力和良好的沟通协作能力; 2、精通Javascript/Java/python编
更新于 2026-06-02杭州