快手数据研发工程师/专家(商业化)-【数据平台】
社招全职1-3年J0012地点:北京 | 杭州状态:招聘
工作描述
任职要求 1、有Hive,Kafka,Spark,Storm,Hbase,Flink等两种以上两年以上使用经验; 2、熟悉数据仓库建设方法和ETL相关技术,对于数据的设计有自己的思考,具备优秀的数学思维和建模思维; 3、熟练使用SQL,对类SQL有过优化经验,对数据倾斜有深度的理解。了解特征工程常用方法…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
Hive+
[英文] Hive Tutorial
https://www.tutorialspoint.com/hive/index.htm
Hive is a data warehouse infrastructure tool to process structured data in Hadoop. It resides on top of Hadoop to summarize Big Data, and makes querying and analyzing easy.
https://www.youtube.com/watch?v=D4HqQ8-Ja9Y
Kafka+
https://developer.confluent.io/what-is-apache-kafka/
https://www.youtube.com/watch?v=CU44hKLMg7k
https://www.youtube.com/watch?v=j4bqyAMMb7o&list=PLa7VYi0yPIH0KbnJQcMv5N9iW8HkZHztH
In this Apache Kafka fundamentals course, we introduce you to the basic Apache Kafka elements and APIs, as well as the broader Kafka ecosystem.
Spark+
[英文] Learning Spark Book
https://pages.databricks.com/rs/094-YMS-629/images/LearningSpark2.0.pdf
This new edition has been updated to reflect Apache Spark’s evolution through Spark 2.x and Spark 3.0, including its expanded ecosystem of built-in and external data sources, machine learning, and streaming technologies with which Spark is tightly integrated.
Apache Storm+
[英文] Tutorial
https://storm.apache.org/releases/2.6.0/Tutorial.html
In this tutorial, you'll learn how to create Storm topologies and deploy them to a Storm cluster.
https://www.baeldung.com/apache-storm
This tutorial will be an introduction to Apache Storm, a distributed real-time computation system.
HBase+
[英文] HBase Tutorial
https://www.tutorialspoint.com/hbase/index.htm
HBase is a data model that is similar to Google's big table designed to provide quick random access to huge amounts of structured data. This tutorial provides an introduction to HBase, the procedures to set up HBase on Hadoop File Systems, and ways to interact with HBase shell.
Flink+
https://nightlies.apache.org/flink/flink-docs-release-2.0/docs/learn-flink/overview/
This training presents an introduction to Apache Flink that includes just enough to get you started writing scalable streaming ETL, analytics, and event-driven applications, while leaving out a lot of (ultimately important) details.
https://www.youtube.com/watch?v=WajYe9iA2Uk&list=PLa7VYi0yPIH2GTo3vRtX8w9tgNTTyYSux
Today’s businesses are increasingly software-defined, and their business processes are being automated. Whether it’s orders and shipments, or downloads and clicks, business events can always be streamed. Flink can be used to manipulate, process, and react to these streaming events as they occur.
还有更多 •••
相关职位
社招5-10年J0012
1、本科及以上学历,拥有5-10年工作经验,对Java、Spring、常用中间件,有深刻的理解与应用; 2、扎实的领域建模、业务架构、DDD经验,平台型工程架构经验; 3、有海量数据处理经验者优先考虑
更新于 2026-02-10北京
社招3年以上技术类
1. 三年以上Java开发经验,熟悉Spring/SpringBoot框架,具备部署高可用的web服务能力; 2. 掌握关系数据库及 SQL 相关知识,熟悉基本的设计和优化原则; 3. 有较强逻辑思考
更新于 2026-06-16上海
社招3-5年J0012
1、本科及以上学历,计算机相关专业,存储领域 3 年以上工作经验; 2、熟练掌握 Java,具备优秀的工程能力,对代码质量有极高要求; 3、熟悉 HDFS、Ceph、S3 等主流分布式存储系统,有实际
更新于 2026-05-19北京|杭州|深圳