
MomentaSRE运维工程师
社招全职5年以上运维地点:北京 | 苏州 | 上海状态:招聘
任职要求
1. 计算机相关专业,五年以上工作经验; 2. 熟悉Linux操作系统,了解Linux操作系统基本原理; 3. 熟悉Elk、Prometheus、Grafana等监控日志工具使用; 4. 熟悉虚拟化和容器技术,如Esxi、Docker、Kubernetes,了解其原理并能够熟练配置; 5. 熟悉 至少一种主流C…
登录查看完整任职要求
微信扫码,1秒登录
工作职责
1. 负责各产品线服务的稳定、高效运行,跟踪用户体验,优化运维架构; 2. 推动运维自动化/标准化方案设计,自动化工具及平台研发,提升运维效率; 3. 及时响应各类故障报警,快速解决问题恢复业务;
包括英文材料
Linux+
https://ryanstutorials.net/linuxtutorial/
Ok, so you want to learn how to use the Bash command line interface (terminal) on Unix/Linux.
https://ubuntu.com/tutorials/command-line-for-beginners
The Linux command line is a text interface to your computer.
https://www.youtube.com/watch?v=6WatcfENsOU
In this Linux crash course, you will learn the fundamental skills and tools you need to become a proficient Linux system administrator.
https://www.youtube.com/watch?v=v392lEyM29A
Never fear the command line again, make it fear you.
https://www.youtube.com/watch?v=ZtqBQ68cfJc
ELK+
https://logz.io/learn/complete-guide-elk-stack/
With millions of downloads for its various components since first being introduced, the ELK Stack is the world’s most popular log management platform.
https://www.baeldung.com/ops/elk
In this tutorial, we’ll learn about the basics of the ELK stack.
https://www.youtube.com/watch?v=jk4RoEYCZTo
explains how to install and configure ELK (Elastic Search, Logstash, Kibana) Stack, a log management solution for analyzing and visualizing your data.
Prometheus+
https://grafana.com/docs/grafana/latest/getting-started/get-started-grafana-prometheus/
Prometheus is an open source monitoring system for which Grafana provides out-of-the-box support.
https://prometheus.io/docs/tutorials/getting_started/
Prometheus is a system monitoring and alerting system.
Grafana+
Docker+
https://www.youtube.com/watch?v=GFgJkfScVNU
Master Docker in one course; learn about images and containers on Docker Hub, running multiple containers with Docker Compose, automating workflows with Docker Compose Watch, and much more. 🐳
https://www.youtube.com/watch?v=kTp5xUtcalw
Learn how to use Docker and Kubernetes in this complete hand-on course for beginners.
Kubernetes+
https://kubernetes.io/docs/tutorials/kubernetes-basics/
This tutorial provides a walkthrough of the basics of the Kubernetes cluster orchestration system.
https://kubernetes.io/zh-cn/docs/tutorials/kubernetes-basics/
本教程介绍 Kubernetes 集群编排系统的基础知识。每个模块包含关于 Kubernetes 主要特性和概念的一些背景信息,还包括一个在线教程供你学习。
https://www.youtube.com/watch?v=s_o8dwzRlu4
Hands-On Kubernetes Tutorial | Learn Kubernetes in 1 Hour - Kubernetes Course for Beginners
https://www.youtube.com/watch?v=X48VuDVv0do
Full Kubernetes Tutorial | Kubernetes Course | Hands-on course with a lot of demos
CI+
https://www.ibm.com/cn-zh/think/topics/continuous-integration
持续集成 (CI) 是一种软件开发实践,开发人员在整个开发周期中会定期将新的代码和代码变更集成到中央代码存储库中。它是 DevOps 和敏捷方法的关键组成部分。
https://www.youtube.com/watch?v=42UP1fxi2SY
CD+
https://www.redhat.com/zh-cn/topics/devops/what-is-ci-cd
CI/CD 是持续集成和持续交付/部署的缩写,旨在简化并加快软件开发生命周期。
https://www.youtube.com/watch?v=R8_veQiYBjI&list=PLy7NrYWoggjzSIlwxeBbcgfAdYoxCIrM2
还有更多 •••
相关职位

社招5年以上
岗位描述: 1、负责同程旅行核心服务的高可靠、稳定、高效运行,持续保障核心业务 SLA 达成; 2、主导服务架构设计与评审,推动 AI 辅助的容量规划、性能调优与智能配置治理; 3、推进系统稳定性与容灾方案设计及落地,结合混沌工程与 AI 预测能力构建自适应韧性架构; 4、主导重大故障应急响应,利用智能运维(AIOps)技术,基于多维度监控、日志、Trace数据进行智能告警、自动止损与根因定位; 5、参与日常值班与线上应急响应,并通过自动化与 AI 工具提升故障处理效率与精准性。
更新于 2026-06-30苏州|北京
社招3年以上程序&技术类
1.负责企业级CICD、服务运行时中间件的可靠性保障与稳定性建设,持续关注系统运行成本,推动降本增效; 2.负责DevOps平台的开发和维护,提供包括但不限于监控、警报、日志记录、部署等基础功能; 3.完善相关应用的监控告警、降级与预案建设,组织故障演练、应急止损、事故复盘等稳定性工作; 4.负责标准自动化DevOps Pipeline的建立和优化,优化运维流程
上海
社招3年以上程序&技术类
1.负责企业人事、财务、数据系统的运维工作,持续关注系统运行成本,推动降本增效; 2.负责审查架构合理性,梳理、识别应用架构风险,解决或推动业务研发解决架构风险; 3.完善相关应用的监控告警、降级与预案建设,组织故障演练、应急止损、事故复盘等稳定性工作;
上海