小红书数据计算引擎研发工程师-搜广推
社招全职3-5年引擎地点:北京 | 上海 | 深圳状态:招聘♡ 收藏
工作描述
任职要求 任职资格 计算机及相关专业本科及以上学历,具备扎实的数据工程或分布式系统开发基础。 熟悉 Flink、Spark、Ray 等主流分布式计算引擎,具备良好的大规模数据处理与系统开发经验。 掌握分布式存储原理,对 Apache Fluss 等新一代流式存储、Paimon 湖仓格式等有认知或实践。 具备良好的系统设计、抽象能力与编码功底,能够驾驭海量数据、高吞吐、复杂生产场景。 具备优秀的驱动力、团队协作与沟通能力,乐于应对复杂技术挑战。 加分项 具备搜索、推荐、广告核心特征生产、大样本处理、索引实时…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
数据结构+
https://www.youtube.com/watch?v=8hly31xKli0
In this course you will learn about algorithms and data structures, two of the fundamental topics in computer science.
https://www.youtube.com/watch?v=B31LgI4Y4DQ
Learn about data structures in this comprehensive course. We will be implementing these data structures in C or C++.
https://www.youtube.com/watch?v=CBYHwZcbD-s
Data Structures and Algorithms full course tutorial java
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
Hadoop+
https://www.runoob.com/w3cnote/hadoop-tutorial.html
Hadoop 为庞大的计算机集群提供可靠的、可伸缩的应用层计算和存储支持,它允许使用简单的编程模型跨计算机群集分布式处理大型数据集,并且支持在单台计算机到几千台计算机之间进行扩展。
[英文] Hadoop Tutorial
https://www.tutorialspoint.com/hadoop/index.htm
Hadoop is an open-source framework that allows to store and process big data in a distributed environment across clusters of computers using simple programming models.
Spark+
[英文] Learning Spark Book
https://pages.databricks.com/rs/094-YMS-629/images/LearningSpark2.0.pdf
This new edition has been updated to reflect Apache Spark’s evolution through Spark 2.x and Spark 3.0, including its expanded ecosystem of built-in and external data sources, machine learning, and streaming technologies with which Spark is tightly integrated.
Hive+
[英文] Hive Tutorial
https://www.tutorialspoint.com/hive/index.htm
Hive is a data warehouse infrastructure tool to process structured data in Hadoop. It resides on top of Hadoop to summarize Big Data, and makes querying and analyzing easy.
https://www.youtube.com/watch?v=D4HqQ8-Ja9Y
Flink+
https://nightlies.apache.org/flink/flink-docs-release-2.0/docs/learn-flink/overview/
This training presents an introduction to Apache Flink that includes just enough to get you started writing scalable streaming ETL, analytics, and event-driven applications, while leaving out a lot of (ultimately important) details.
https://www.youtube.com/watch?v=WajYe9iA2Uk&list=PLa7VYi0yPIH2GTo3vRtX8w9tgNTTyYSux
Today’s businesses are increasingly software-defined, and their business processes are being automated. Whether it’s orders and shipments, or downloads and clicks, business events can always be streamed. Flink can be used to manipulate, process, and react to these streaming events as they occur.
推荐系统+
[英文] Recommender Systems
https://www.d2l.ai/chapter_recommender-systems/index.html
Recommender systems are widely employed in industry and are ubiquitous in our daily lives.
信息检索+
https://nlp.stanford.edu/IR-book/information-retrieval-book.html
Christopher D. Manning, Prabhakar Raghavan and Hinrich Schütze, Introduction to Information Retrieval, Cambridge University Press. 2008.
还有更多 •••
相关职位
社招1-3年J0012
1、本科及以上学历,计算机科学与技术、软件工程或相关专业方向; 2、熟悉 Java 语言,扎实的计算机基础; 3、熟悉至少一种主流大数据引擎,包括但不限于 Spark/Presto/Flink/Kyl
更新于 2026-06-22北京|杭州
社招1年以上
1.熟练掌握 Java,具备良好的编码能力和工程实践经验。 2.具备较强的工程意识,关注代码质量、系统稳定性、可维护性、资源成本和交付效率。 3.善于进行性能分析与优化,能够通过指标、日志、火焰图等手
更新于 2026-08-25北京
社招3年以上技术-基础平台
1、熟练掌握Java,熟悉各种编译、调试、性能分析工具,具备扎实的编码能力和良好的代码习惯。 2、对Flink和Calcite有深入研究与开发经验者优先。 3、具备扎实的工程基础,有大规模分布式应用开
更新于 2026-08-03北京