小红书数据引擎AIOps/Agent专家
社招全职3-5年数据引擎地点:北京 | 上海 | 杭州状态:招聘
工作描述
任职要求 1、计算机相关专业。 2、熟悉一种或多种大数据引擎,如Flink、Spark、Ray、Kafka、ClickHouse、Doris等。负责过大规模的数据体量场景,并能结合引擎原理和业务场景进行优化和治理。 3、熟悉并落地如下1个或多个AIOps、Agent领域的经验: a). 大规模云平台的资源分配、调度优化和中长期资源规划:运用需求预测、运筹优化等方法,解决大规模混部环境的资源管理和技术风险问题,突破传统策略的瓶颈。 b)…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大数据+
https://www.youtube.com/watch?v=bAyrObl7TYE
https://www.youtube.com/watch?v=H4bf_uuMC-g
With all this talk of Big Data, we got Rebecca Tickle to explain just what makes data into Big Data.
Flink+
https://nightlies.apache.org/flink/flink-docs-release-2.0/docs/learn-flink/overview/
This training presents an introduction to Apache Flink that includes just enough to get you started writing scalable streaming ETL, analytics, and event-driven applications, while leaving out a lot of (ultimately important) details.
https://www.youtube.com/watch?v=WajYe9iA2Uk&list=PLa7VYi0yPIH2GTo3vRtX8w9tgNTTyYSux
Today’s businesses are increasingly software-defined, and their business processes are being automated. Whether it’s orders and shipments, or downloads and clicks, business events can always be streamed. Flink can be used to manipulate, process, and react to these streaming events as they occur.
Spark+
[英文] Learning Spark Book
https://pages.databricks.com/rs/094-YMS-629/images/LearningSpark2.0.pdf
This new edition has been updated to reflect Apache Spark’s evolution through Spark 2.x and Spark 3.0, including its expanded ecosystem of built-in and external data sources, machine learning, and streaming technologies with which Spark is tightly integrated.
Ray+
https://github.com/ray-project/ray
Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
https://www.youtube.com/watch?v=FhXfEXUUQp0
In this video, I'll teach you everything you need to know about Apache Ray!
https://www.youtube.com/watch?v=fMiAyj2kgac
Using powerful machine learning algorithms is easy using Ray.io and Python.
https://www.youtube.com/watch?v=q_aTbb7XeL4
Parallel and Distributed computing sounds scary until you try this fantastic Python library.
Kafka+
https://developer.confluent.io/what-is-apache-kafka/
https://www.youtube.com/watch?v=CU44hKLMg7k
https://www.youtube.com/watch?v=j4bqyAMMb7o&list=PLa7VYi0yPIH0KbnJQcMv5N9iW8HkZHztH
In this Apache Kafka fundamentals course, we introduce you to the basic Apache Kafka elements and APIs, as well as the broader Kafka ecosystem.
ClickHouse+
[英文] Advanced Tutorial
https://clickhouse.com/docs/tutorial
Learn how to ingest and query data in ClickHouse using the New York City taxi example dataset.
https://www.youtube.com/watch?v=FtoWGT7kS-c
ClickHouse is an open-source column-oriented DBMS for online analytical processing that allows users to generate analytical reports using SQL queries in real-time.
https://www.youtube.com/watch?v=Rhe-kUyrFUE&list=PL0Z2YDlm0b3gcY5R_MUo4fT5bPqUQ66ep
还有更多 •••
相关职位
实习工程-前端类
1、本科及以上学历,计算机相关专业; 2、至少熟悉一种主流前端MVVM框架,Vue、React等; 3、熟练掌握HTML5、CSS3、ES6等,熟悉HTTP协议和Node.js开发; 4、计算机基础扎
更新于 2025-11-25北京
社招1-3年J0012
1、本科及以上学历,计算机科学与技术、软件工程或相关专业方向; 2、熟悉 Java 语言,扎实的计算机基础; 3、熟悉至少一种主流大数据引擎,包括但不限于 Spark/Presto/Flink/Kyl
更新于 2026-06-22北京|杭州
社招3-5年数据引擎
1. 本科及以上学历,3年以上AI&Data引擎/数据/存储研发经验 2. 加分项:熟悉大模型技术和产品生态,如Data-Juicer/Ray/Daft/Pytorch/RAG等 3. 熟悉Pytho
更新于 2026-08-01北京|上海|杭州