站点可靠性工程师,基础设施平台 — 英国(中级至高级资深)
查看雇主原标题
Site Reliability Engineer, Infrastructure Platforms — UK (Intermediate to Senior Staff)GitLab · Remote
职位信息来自雇主公开的招聘页面。申请前请务必在雇主官网核实详情。
为什么值得关注?
发现指数 50/100,仅依据与该职位一起存储的证据计算。
- 新的雇主官方职位
- 远程职位
分数构成
- 时效性 (随职位发布时间变化)+18
- 雇主官方来源+15
- 远程职位+8
- 稀有职位+1
- 公司来源健康度+8
该职位未包含:已披露薪资、提及签证担保、提及搬迁、未出现在监控的职位板上。
这些理由来自雇主自己的职位描述与我们核实过的来源检查结果。除了已存储的信号之外,我们不做任何推测。
职位描述
机器翻译GitLab 是面向 DevSecOps 的智能编排平台。GitLab 帮助组织提升开发者生产力、提高运营效率、降低安全与合规风险,并加速数字化转型。超过 5000 万注册用户以及超过 50% 的《财富》100 强企业*都信赖 GitLab,以更快交付更优质、更安全的软件。
我们产品中内置的相同原则也体现在我们团队的工作方式中:我们将 AI 视为核心生产力倍增器,所有团队成员都应把 AI 融入日常工作流程,以推动效率、创新和影响力。GitLab 是职业加速发展、创新蓬勃生长、每个声音都受到重视的地方。我们的高绩效文化由我们的价值观和持续的知识交流驱动,使团队成员能够充分发挥潜力,同时与行业领袖协作解决复杂问题。与我们一起共创未来,打造改变世界软件开发方式的技术。
* Fortune 500® 是 Fortune Media IP Limited 的注册商标,经许可使用。该声明基于 GitLab 数据。Fortune 100 指 2025 年 6 月发布的 2025 年 Fortune 500 榜单中排名前 20% 的公司。Fortune 和 Fortune Media IP Limited 与 GitLab 无关联,也不为 GitLab 的产品或服务背书。
职位概览
站点可靠性工程师负责让 GitLab 面向用户的服务和生产系统在大规模环境下可靠运行。他们将软件工程与卓越运营相结合,运用良好的工程原则、自动化和持续改进来构建、运营并演进我们的生产基础设施。
这是一个针对基础设施平台多个站点可靠性工程机会的统一申请。我们不会要求你预先选择正确的团队或级别,而是会全面评估你的技能,并将你匹配到最符合你的经验和我们的招聘需求的机会。我们正在为多个基础设施平台团队招聘从中级到高级资深(Senior Staff)的站点可靠性工程师。
我们并不期望每位候选人都具备我们环境中每一项技术的经验。我们寻找的是具备扎实技术基础、成长型思维以及快速学习能力的工程师。我们会支持你熟练掌握 GitLab 的工具、系统和工作方式。
请注意:该职位仅面向位于英国的候选人开放。位于美国或加拿大的候选人可以申请此职位:站点可靠性工程师,基础设施平台 — AMER(中级至高级资深)
我们的 SRE 招聘流程如何运作
由于这是一个针对基础设施平台多个 SRE 职位的统一申请,我们的流程旨在对你进行一次评估并做好匹配,而不是为每个团队分别面试。
• 招聘人员初筛:围绕你的背景、你正在寻找的机会,以及适合的级别和团队进行交流,以便我们为你的流程指明正确方向。
• 核心技术面试:所有 SRE 候选人都会参加的共享评估,无论最终加入哪个团队。这是一场低压力、协作式的讨论,涵盖系统架构和事故复盘。
• 招聘经理面试:围绕主人翁意识、判断力、执行力、协作和成长进行交流,这些是让 SRE 在 GitLab 高效发挥作用的非技术信号。
• 同级技术面试:由你最有可能加入的团队中的 SRE 主持的团队特定深入面试,聚焦该团队实际处理的问题。
• 越级面试:与高级领导者交流价值观契合度,以及你将如何跨团队协作。
面试结束后,我们会将你的表现与我们当前的招聘需求一并考虑,以确认你最能发挥所长的级别和团队。面试结果是主要因素,最终安排也会反映我们当时的实际招聘优先级。
我们会在整个面试流程中,根据你经验的范围和影响力来校准你的级别。
• 中级:你能够在明确界定的领域内独立交付有意义的可靠性改进。
• 高级:你能够端到端负责复杂的可靠性工作,并提升团队的有效性。
• 资深:你能够影响多个团队的可靠性,解决系统性问题,并创建他人可复用的方法。
• 高级资深:你能够为更广泛的基础设施领域设定技术方向,并在组织规模上影响可靠性战略。
岗位职责
• 保持面向用户的服务和生产系统可靠、可扩展且高效
• 构建自动化和工具,减少繁琐工作,并用可重复、基础设施即代码驱动的工作流取代人工操作
• 在 Kubernetes 上运营和排查生产系统,包括部署、发布和扩缩容
• 编写和维护基础设施即代码,并通过 CI/CD 和 GitOps 安全发布变更
• 参与值班、分诊告警、遵循并改进运行手册,并在适当时升级处理
• 为可观测性技术栈做出贡献,使用指标、日志和 SLO 及早发现症状,而不仅仅是发现故障
• 参与事故响应和事故后复盘,将经验教训转化为自动化和流程变更
• 记录运行手册、架构决策和复盘,让你的发现成为可重复的实践
你将带来什么
• 保持生产系统可靠运行的经验,兼具运营思维和真实的软件工程实践
• 构建全新基础设施工具和自动化的经验,而不仅仅是配置现有工具。例如,Terraform 模块、Kubernetes operator 或 controller,或从零编写的生产自动化和服务
• 阅读、调试和理解代码的能力。我们的大多数团队使用 Go;一些团队使用 Ruby。你能够讨论一段代码的行为、性能和故障模式
• 具备基础设施即代码的经验,以及与你级别相称的 Kubernetes 及其生态系统的深度经验
• 至少一家主要云服务提供商(GCP 或 AWS)的实操经验
• 熟悉可观测性实践,包括指标、日志、告警和 SLO 或 SLI,并使用数据为运营决策提供依据
• 能够从容参与值班和事故响应,并在压力下以结构化方法进行故障排查
• 出色的书面沟通能力,以及在异步、分布式环境中作为“一人管理者”运作的能力
• 有使用自动化,并越来越多地使用 AI 来减少繁琐工作并改善你和团队工作方式的记录
• 与 GitLab 的价值观一致,并承诺按照这些价值观开展工作
福利待遇
• 灵活带薪休假
• 团队成员资源小组
• 股权薪酬与员工购股计划
• 成长与发展基金
• 育儿假
请注意,我们欢迎具有不同经验水平的候选人表达兴趣;许多成功候选人并不满足每一项要求。此外,研究表明,来自代表性不足群体的人除非满足每一项资格要求,否则不太可能申请工作。如果你对这个职位感到兴奋,请申请,并让我们的招聘人员评估你的申请。
国家招聘准则:GitLab 在世界各国招聘新团队成员。我们的所有职位均为远程,但某些职位可能有特定的基于地点的资格要求。我们的招聘团队可以在启动招聘流程后帮助回答任何关于地点的问题。
隐私政策:请查看我们的招聘隐私政策。你的隐私对我们很重要。
GitLab 自豪地成为提供平等机会的工作场所,并且是积极行动雇主。GitLab 在招聘、雇佣、职业发展和晋升、升职以及退休方面的政策和实践完全基于能力,不论种族、肤色、宗教、血统、性别(包括怀孕、哺乳、性取向、性别认同或性别表达)、国籍、年龄、公民身份、婚姻状况、精神或身体残疾、遗传信息(包括家族病史)、退伍状态、受保护退伍军人身份(包括残疾退伍军人、近期退伍军人、战时或战役徽章现役退伍军人,以及武装部队服务奖章退伍军人),或任何其他受法律保护的基础。GitLab 不会容忍基于任何这些特征的歧视或骚扰。另请参阅 GitLab 的 EEO 政策和 EEO is the Law。如果你有残疾或需要便利支持的特殊需求,请在招聘流程中告知我们。
以上内容由机器翻译自动生成,可能存在错误;投递前请以雇主原文为准。
查看雇主原文
职位描述
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster.
The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software.
* Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab.
An overview of this role
Site Reliability Engineers keep GitLab's user-facing services and production systems running reliably at scale. They combine software engineering with operational excellence, applying sound engineering principles, automation, and continuous improvement to build, operate, and evolve our production infrastructure.
This is a single application for Site Reliability Engineering opportunities across Infrastructure Platforms. Rather than asking you to choose the right team or level upfront, we evaluate your skills holistically and match you to the opportunity that best aligns with your experience and our hiring needs. We hire Site Reliability Engineers from Intermediate through Senior Staff across multiple Infrastructure Platforms teams.
We don't expect every candidate to have experience with every technology in our environment. We're looking for engineers with strong technical fundamentals, a growth mindset, and the ability to learn quickly. We'll support you in becoming successful with GitLab's tools, systems, and ways of working.
Please note: This position is open to candidates based in the United Kingdom only. Candidates based in the United States or Canada can apply to this posting: Site Reliability Engineer, Infrastructure Platforms — AMER (Intermediate to Senior Staff)
How our SRE hiring works
Because this is a single application for SRE roles across Infrastructure Platforms, our process is built to evaluate you once and match you well, rather than interviewing separately for every team.
• Recruiter Screen: A conversation about your background, what you're looking for, and the level and teams that fit, so we can point your process in the right direction.
• Core Technical: The shared assessment every SRE candidate takes, regardless of eventual team. A low-stress, collaborative discussion covering system architecture and incident review.
• Hiring Manager Interview: A conversation about ownership, judgment, execution, collaboration, and growth, the non-technical signals that make an SRE effective at GitLab.
• Peer Technical: Team-specific depth, run by SREs from the team you're most likely to join, focused on the problems that team actually works on.
• Skip-Level Interview: A conversation with a senior leader on values alignment, and how you'll work across teams.
After your interviews, we consider your performance alongside our current hiring needs to confirm the level and team where you'll do your best work. Interview results are a major factor, and final placement also reflects our active hiring priorities at the time.
We’ll calibrate your level throughout the interview process based on the scope and impact of your experience.
• Intermediate: You independently deliver meaningful reliability improvements within a defined area.
• Senior: You own complex reliability work end to end and raise the effectiveness of your team.
• Staff: You shape reliability across multiple teams, solving systemic problems and creating approaches others can reuse.
• Senior Staff: You set technical direction across a broader Infrastructure area and influence reliability strategy at organizational scale.
岗位职责
• Keep user-facing services and production systems reliable, scalable, and efficient
• Build automation and tooling that reduces toil and replaces manual work with repeatable, infrastructure-as-code-driven workflows
• Operate and troubleshoot production systems on Kubernetes, including deployments, rollouts, and scaling
• Write and maintain infrastructure as code, and ship changes safely through CI/CD and GitOps
• Participate in on-call, triage alerts, follow and improve runbooks, and escalate appropriately
• Contribute to the observability stack, using metrics, logs, and SLOs to detect symptoms early rather than just outages
• Take part in incident response and post-incident reviews, turning learnings into changes in automation and process
• Document runbooks, architecture decisions, and reviews so your findings become repeatable practices
What you'll bring
• Experience keeping production systems reliable, combining an operations mindset with real software engineering practice
• Experience building net-new infrastructure tooling and automation, not just configuring existing tools. For example, Terraform modules, Kubernetes operators or controllers, or production automation and services written from scratch
• The ability to read, debug, and reason about code. Most of our teams work in Go; some work in Ruby. You can discuss a piece of code's behavior, performance, and failure modes
• Experience with infrastructure as code, and with Kubernetes and its ecosystem, at a depth appropriate to your level
• Hands-on experience with at least one major cloud provider (GCP or AWS)
• Familiarity with observability practices, including metrics, logging, alerting, and SLOs or SLIs, and using data to inform operational decisions
• Comfort participating in on-call and incident response, with a structured approach to troubleshooting under pressure
• Strong written communication and the ability to operate as a manager-of-one in an async, distributed environment
• A track record of using automation, and increasingly AI, to reduce toil and improve how you and your team work
• Alignment with GitLab's values and a commitment to working in accordance with them
福利待遇
• Flexible Paid Time Off
• Team Member Resource Groups
• Equity Compensation & Employee Stock Purchase Plan
• Growth and Development Fund
• Parental Leave
Please note that we welcome interest from candidates with varying levels of experience; many successful candidates do not meet every single requirement. Additionally, studies have shown that people from underrepresented groups are less likely to apply to a job unless they meet every single qualification. If you're excited about this role, please apply and allow our recruiters to assess your application.
Country Hiring Guidelines: GitLab hires new team members in countries around the world. All of our roles are remote, however some roles may carry specific location-based eligibility requirements. Our Talent Acquisition team can help answer any questions about location after starting the recruiting process.
Privacy Policy: Please review our Recruitment Privacy Policy. Your privacy is important to us.
GitLab is proud to be an equal opportunity workplace and is an affirmative action employer. GitLab’s policies and practices relating to recruitment, employment, career development and advancement, promotion, and retirement are based solely on merit, regardless of race, color, religion, ancestry, sex (including pregnancy, lactation, sexual orientation, gender identity, or gender expression), national origin, age, citizenship, marital status, mental or physical disability, genetic information (including family medical history), discharge status from the military, protected veteran status (which includes disabled veterans, recently separated veterans, active duty wartime or campaign badge veterans, and Armed Forces service medal veterans), or any other basis protected by law. GitLab will not tolerate discrimination or harassment based on any of these characteristics. See also GitLab’s EEO Policy and EEO is the Law . If you have a disability or special need that requires accommodation , please let us know during the recruiting process .