字节跳动Intelligent Computing Architecture and Operating System Researcher | 智能计算体系结构与操作系统研究员-筋斗云人才计划
任职要求
1. Academic Background: Got doctor degree, preferably majoring in software engineering, computer science, mathematics, artificial intelligence, or related fields. Strong capabilities in computer architecture; excellent coding skills, solid foundation in data structures and fundamental algorithms; proficiency in C/C++, Go, or Python. 2. Technical Proficiency Familiar with Linux operating systems, kernel, and network-related knowledge; prior development experience in these areas is preferred. Experience in performance optimization is a plus. 3. Problem-Solving Abilities Outstanding problem analysis and solving skills, with the ability to independently explore…
工作职责
Team Introduction: The ByteDance System Department is responsible for the R&D, design, procurement, delivery, and operational management of the company's infrastructure ranging from chips to servers, operating systems, networks, CDNs, and data centers. It provides efficient, stable, and scalable infrastructure to support global services such as Douyin, Toutiao, and Volcano Engine. The current areas of operation include, but are not limited to: the design and construction of data centers, chip R&D, server development, network engineering, Volcano Engine's edge-cloud services, high-performance intelligent hardware development, intelligent delivery and operation of IDC resources, intelligent monitoring and early warning of hardware infrastructure, operating systems and kernels, virtualization technologies, compilation toolchains, supply chain management, and many other infrastructure-related areas. 团队介绍: 字节跳动系统部,负责字节跳动从芯片到服务器、操作系统、网络、CDN 、数据中心等基础设施的研发、设计、采购、交付与运营管理,为包含抖音、头条、火山引擎等全球业务提供高效、稳定、具备可扩展性的基础设施。部门当前业务开展包括不限于:数据中心设计建设、芯片研发、服务器研发、网络工程研发、火山引擎边缘云业务、高性能智能硬件研发、IDC资源智能交付与运维、硬件基础设施智能监控与预警、操作系统与内核、虚拟化技术、编译工具链、供应链管理等众多基础设施相关方向。 课题介绍: 在当今数字化时代,随着云计算、人工智能和大数据技术的深度融合,现代数据中心正面临着指数级增长的算力需求与现有计算架构效能瓶颈之间的突出矛盾。传统以通用CPU为核心的体系架构在应对多样化负载时,暴露出诸多问题。例如,内存子系统带宽与时延约束导致的 “内存墙” 效应持续加剧,异构计算单元间的数据搬运开销占比超过实际运算时间,安全可信执行环境带来的性能损耗超过 30%,单机柜算力密度提升受限于功耗密度阈值。与此同时,新兴工作负载如AI训练、图计算、时序数据库等呈现出动态异构特征,对计算架构提出了差异化需求,传统固定架构难以实现最优能效比。 操作系统作为计算机体系结构下重要的软件基础设施与核心技术,在这样的背景下也面临着巨大的挑战。随着计算需求的增长和技术的进步,传统的同构计算环境已无法满足日益复杂的计算任务。现代计算场景中,硬件架构呈现高度异构化,包括 CPU、GPU、FPGA、TPU、NPU、DPU 等,同时边缘计算、云计算形成分布式网络。传统操作系统难以高效管理跨节点、跨架构的资源。加之人工智能训练等场景需要低延迟、高吞吐、安全可信,动态弹性的分布式系统支持,这就要求操作系统具备跨异构资源的统一抽象与调度能力。学术界和工业界对下一代计算机操作系统在分布式微内核架构,异构资源调度算法,跨层优化与编译器支持,安全可信技术,虚拟化和 Serverless,AI 驱动操作系统内核优化以及操作系统内置 AI 推理引擎等方面展开了积极的探索和研究。 课题挑战: 方向一:体系化结构方向 1)负载特征与架构优化:建立数据中心动态负载特征建模框架,深入研究面向数据中心Workload的体系结构设计与优化方法,使系统能够更好地适应多样化的负载需求; 2)CPU核心架构创新:研究高性能低功耗CPU核心架构,积极探索超标量流水线与数据流引擎的融合设计,提升CPU的性能和能效; 3)新型内存层次构建:构建支持存算一体化的新型内存层次结构,研究基于3D堆叠技术的近存计算架构,重点突破高带宽互连拓扑优化、混合内存控制器设计、内存访问模式预测算法,解决 “内存墙” 等问题; 4)安全可信架构构建:构建安全可信计算架构,包括侧信道攻击防御的微架构级实现、侧信道安全架构、自动侧 / 隐蔽通道泄漏检测,确保系统在复杂环境下的安全性和完整性; 5)数据中心架构创新:探索整机柜级系统总线扩展,构建内存语义互联的新型数据中心架构,研究基于新型总线协议 (CXL/UALink) 的全局内存共享机制,提升数据中心的整体性能和资源利用率; 6)可靠性增强技术研究:研究可靠性增强技术,包括开发基于机器学习的故障预测模型,设计自修复的微架构容错机制,研究硬件静默故障检测,以及系统及IP可靠性特性研究和数据分析,保障系统的稳定运行。 方向二:操作系统方向 1)操作系统关键技术突破:突破传统单机操作系统存在的硬件资源利用局限、功能扩展与升级运维复杂、数据管理与共享不足、安全性与可靠性欠佳等问题。在计算高度异构以及计算环境分布化的情况下,从硬件到软件建立完整的信任链,保证整个系统的安全性和完整性。同时,有效地管理和协调多个节点间的通信、数据同步及故障恢复,设计高效的调度算法来匹配任务需求与最适合的计算资源,以最大化性能和效率。操作系统需要能够理解不同类型的计算任务,并能根据实时的工作负载动态调整资源分配,实现跨异构资源的统一抽象与调度; 2)跨领域知识融合:本课题需要融合OS、内核、算法、存储、虚拟化、网络、系统工程等多方面的跨领域知识和经验,以实现数据中心智能计算体系结构与操作系统的协同创新。
We are aiming to leverage AI and other leading technology and dedicated to provide safe and reliable risk control capabilities behind payments. The core technologies include rule engines, model engines, intelligent algorithm models, etc., We are the leading platform with capabilities of high concurrent real-time risk calculations and massive big data analysis and processing. And as the core risk management tech platform for global payment business, we adopt a multi-center deployment architecture around the world. Here you may have the opportunity to learn more about and participate in the design and development of the following aspects: 1. Ultimate computing optimization at the millisecond level. 2. Behavior analysis and risk mining under massive data. 3. Global multi-center system architecture planning and high-availability solution design. 4. Participated in the design of R&D of risk control systems and big data platforms. You will also have the opportunity to explore the architectural design and implementation of cutting-edge technologies such as privacy computing and large models in risk control systems.
阿里巴巴国际数字商业 (AIDC) 商业智能部招聘 公司&部门介绍 阿里国际数字商业集团(AIDC),是阿里巴巴集团的核心且快速增长的业务。 契合国家“一带一路”的战略方向,在国际化的大蓝海赛道上高速驰骋,连续多年收入增长在30%以上。旗下业务覆盖近200个国家、市场,服务于4亿的全球消费者,囊括跨境电商、本地电商、B2B、O2O零售、供应链网络等多元化的业务形态,在全球近30个国家与地区设置办公地点,拥有超过20种不同国籍的员工。 商业智能部是AIDC的商业分析与决策支持部门,我们依托于AIDC的全球化大数据以及阿里多年的商业分析经验沉淀,拥有来自于互联网、咨询、投资、传统行业等多元化背景的团队,产生多视角、全方位的立体洞察,支持集团以及各事业部与职能部门管理者的关键决策。 我们深入业务,通过有智慧的数据洞察, 紧密结合的业务场景,实现从宏观市场到微观战术的商业分析与判断,为决策层提供关键数据洞察、核心策略参考、支持主要的各项决策制定与优化,达到数据驱动业务的完美实践。 Team and Role Introduction: Working under AIDC BI family, this role is BI analyst supporting E-commerce Platform Miravia. Responsibilities: Key tasks and responsibilities: • Collect, clean, and maintain user data from various sources including marketing platforms, mobile app, web analytics etc. • Be the expert in using data to measure and analyze business performance in each our markets and lines of business. • Collaborate with product managers, marketing teams, and other stakeholders to identify areas of opportunity for improving user growth and engagement. • Lead new data analytics capability rollouts and/or data-led initiatives throughout Arise Project • Analyze user behavior data to identify trends, patterns, and insights that can be used to improve user engagement, retention and acquisition. • Design market/business intelligence reports and performance measurement dashboards to share with senior management• Perform ad-hoc business analysis, to drill down on certain business challenges, providing conclusions and advice based on data analysis. • Monitor and evaluate the effectiveness of user growth initiatives and make recommendations for optimization and improvement.
The Assortment, Content & Ads Governance Team (ACAG team) is part of the Lazada’s Risk and Security, and it is charged with the mission of developing a comprehensive strategy for Lazada with regards to assortment, so as to foster a healthy and safe e-commerce environment for our users. The role is responsible for developing and implementing proactive strategies and operational processes to protect users from Assortment and Content related risks (Prohibited and Controlled Goods, IP infringement products, hate speech etc). You will have access to analytical tools to develop and implement strategies and solutions using data driven methodologies to mitigate the risks associated with platform assortment and content. Responsibilities: - Develop a deep understanding of the eCommerce customer and seller journey, including registration and onboarding, product listing, order placement, payment, user interactions, returns and refunds, user reports and feedback, etc. - Develop subject matter expertise on eCommerce platform operation and governance, where rules, strategies, and enforcements are effectively established to ensure users are compliant with platform policies. - Lead cross-functional efforts to enhance platform policies and operational mechanisms, fostering a collaborative environment to support ongoing strategy refinement. - Work with large data sets to analyze patterns, trends, and modus operandi of platform operation and governance issues (as well as merchants who perpetrate these issues). - Make data-driven recommendations on prioritization of controls for platform governance and product compliance. - Collaborate with PD, Tech, and Algo counterparts to build machine learning models and rules to detect assortment and user related operation and governance issues on the platform. - Operate the risk engine, including the creation and continuous evaluation of rules to prevent and detect platform operation and governance issues. - Capture and communicate findings with internal and external stakeholders through dashboards, periodic reports, and presentations.
1) Conduct comprehensive business insights analysis covering user behaviour, trend extraction, data interpretation, and insights generation in Search products. 2) Design and execute analysis to uncover insights critical for key business decisions. 3) Develop and maintain data frameworks, pipelines, and dashboards to identify and mitigate risks, as well as to capitalise on opportunities. 4) Collaborate with cross-functional teams to understand requirements, provide actionable insights, and drive continuous improvements in Search products. 5) Present findings and recommendations to stakeholders in a clear and compelling manner. 1. 进行全面的商业洞察分析,涵盖搜索产品中的用户行为分析、趋势分析、数据解读和洞察生成。 2. 设计和执行分析,揭示对关键业务决策至关重要的洞察。 3. 开发和维护数据框架、数据管道和报表,方便业务团队日常看数,及时发现机会点和风险点。 4. 与跨职能团队合作,了解需求,提供可操作的洞察建议,并推动搜索产品的持续改进。 5. 以清晰、有说服力的方式向决策层汇报分析结果,提供建议。