腾讯智能体-数据分析工程师(数仓方向)-WorkBuddy
社招全职3年以上腾讯云-Codebuddy-C端专用技术地点:深圳状态:招聘♡ 收藏
工作描述
任职要求 1.计算机相关专业本科及以上学历,3年以上大数据开发或数仓建设经验;熟练掌握Hadoop、Spark、Flink等大数据组件,有丰富的实际项目经验;熟悉Hive、HBase、Kafka等常用大数据技术,具备数据建模和ETL开发能力; 2.具备扎实的SQL和编程能力(如Java/Python/Scala),能独立完成复杂的数据处理逻辑; 3.良好的逻辑思维和问题解决能力,能够快速定位和解决数据相关问题;责任心强,具备良好的团队协作能力和沟通能力。 加分项 1.熟悉实时计算和流式处理,有Flink/Spark Streami…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
大数据+
https://www.youtube.com/watch?v=bAyrObl7TYE
https://www.youtube.com/watch?v=H4bf_uuMC-g
With all this talk of Big Data, we got Rebecca Tickle to explain just what makes data into Big Data.
Hadoop+
https://www.runoob.com/w3cnote/hadoop-tutorial.html
Hadoop 为庞大的计算机集群提供可靠的、可伸缩的应用层计算和存储支持,它允许使用简单的编程模型跨计算机群集分布式处理大型数据集,并且支持在单台计算机到几千台计算机之间进行扩展。
[英文] Hadoop Tutorial
https://www.tutorialspoint.com/hadoop/index.htm
Hadoop is an open-source framework that allows to store and process big data in a distributed environment across clusters of computers using simple programming models.
Spark+
[英文] Learning Spark Book
https://pages.databricks.com/rs/094-YMS-629/images/LearningSpark2.0.pdf
This new edition has been updated to reflect Apache Spark’s evolution through Spark 2.x and Spark 3.0, including its expanded ecosystem of built-in and external data sources, machine learning, and streaming technologies with which Spark is tightly integrated.
Flink+
https://nightlies.apache.org/flink/flink-docs-release-2.0/docs/learn-flink/overview/
This training presents an introduction to Apache Flink that includes just enough to get you started writing scalable streaming ETL, analytics, and event-driven applications, while leaving out a lot of (ultimately important) details.
https://www.youtube.com/watch?v=WajYe9iA2Uk&list=PLa7VYi0yPIH2GTo3vRtX8w9tgNTTyYSux
Today’s businesses are increasingly software-defined, and their business processes are being automated. Whether it’s orders and shipments, or downloads and clicks, business events can always be streamed. Flink can be used to manipulate, process, and react to these streaming events as they occur.
Hive+
[英文] Hive Tutorial
https://www.tutorialspoint.com/hive/index.htm
Hive is a data warehouse infrastructure tool to process structured data in Hadoop. It resides on top of Hadoop to summarize Big Data, and makes querying and analyzing easy.
https://www.youtube.com/watch?v=D4HqQ8-Ja9Y
HBase+
[英文] HBase Tutorial
https://www.tutorialspoint.com/hbase/index.htm
HBase is a data model that is similar to Google's big table designed to provide quick random access to huge amounts of structured data. This tutorial provides an introduction to HBase, the procedures to set up HBase on Hadoop File Systems, and ways to interact with HBase shell.
Kafka+
https://developer.confluent.io/what-is-apache-kafka/
https://www.youtube.com/watch?v=CU44hKLMg7k
https://www.youtube.com/watch?v=j4bqyAMMb7o&list=PLa7VYi0yPIH0KbnJQcMv5N9iW8HkZHztH
In this Apache Kafka fundamentals course, we introduce you to the basic Apache Kafka elements and APIs, as well as the broader Kafka ecosystem.
ETL+
https://www.ibm.com/think/topics/etl
ETL—meaning extract, transform, load—is a data integration process that combines, cleans and organizes data from multiple sources into a single, consistent data set for storage in a data warehouse, data lake or other target system.
https://www.youtube.com/watch?v=OW5OgsLpDCQ
It explains what ETL is and what it can do for you to improve your data analysis and productivity.
还有更多 •••
相关职位
社招3-5年
1. 计算机科学、人工智能、软件工程、数据科学或相关专业本科及以上学历。 2 精通Python/Java/Go其中的一种或者多种编程语言。 3. 熟悉大语言模型(LLM)原理及应用,具备Prompt
更新于 2026-01-05深圳
社招3年以上腾讯云-Code
1.统招本科及以上学历,统计学、数学、计算机科学、经济学或运筹学相关专业背景; 2.3年以上互联网数据分析经验,熟练编写复杂SQL查询,掌握Hive/Spark/ClickHouse等大数据处理技术
更新于 2026-09-17北京
实习阿里巴巴研究型实
1)计算机科学、人工智能、软件工程或相关专业在读硕士,博士; 2)有相关领域顶会论文(ICML,ICLR,AAAI,NeuralPS,VLDB,SIGMOD等)发表经验; 3)扎实的工程能力,优良的编
更新于 2026-03-17杭州