donator JOBS

IT 与编程

Senior Site Reliability Engineer(远程职位)

Gradle

限欧洲远程
申请

链接会跳转到原始招聘信息。Donator 不代收申请。

可分享的海报
发布时间: (23 天前)有效期: 截至 2026年10月2日

Senior Site Reliability Engineer:「IT 与编程」类别,完全远程

这条招聘信息来自 Gradle,岗位是 Senior Site Reliability Engineer:这是高级职位,公司期待由你自己拿主意。职位属于「IT 与编程」类别,完全远程。公司招的是居住在欧洲的候选人。

招聘信息单独点名了 Java,也就是说,正是这一项工具的使用经验会成为决定因素。

公司没有公布具体数字,这一项会在面试时谈定。招聘信息写明了 GMT 时区,工作时间要和自己的作息对上。

这个职位通过了自动核查:凡是要求外国工作许可、签证担保、特定国籍,或者必须居住在指定国家的招聘信息,都不会进入列表。

工具

要点

公司
Gradle
类别
IT 与编程
谁可以申请
来自欧洲任何国家
工作方式
完全远程
级别
高级
工具
Java
用工形式
全职
发布时间
2026年8月3日 (23 天前)
有效期
截至 2026年10月2日
来源
Himalayas

新职位邮件提醒:IT 与编程

职位板每天更新数次。只有出现新的、已核查的职位时才会写信。不发垃圾邮件,也不需要注册账号。

已经订阅过了? 管理订阅设置

Telegram 上的最新职位

每一条新的远程职位,发布即推送。无需注册。

打开频道

@donator_jobs_zh

公司发布的职位描述

Who We Are AI is changing how software gets built. Code production is becoming a commodity. The focus is shifting from writing code to orchestrating, verifying, and governing change - and the toolchain is the new constraint. We are at the center of this shift. We build Develocity, a toolchain observability and intelligence platform used by some of the world's leading software organizations - Netflix, Airbnb, Spotify, SAP, major global banks, and hundreds more. Develocity helps software teams achieve delivery excellence through deep observability, build and test acceleration, and AI-powered intelligence across the entire toolchain - with current support for Gradle Build Tool, Apache Maven™, sbt, npm, and Python. We are an AI-native company. AI is not a feature we're bolting on - it's central to how we work, how we think about our product, and where we're heading. We're investing deeply in making Develocity's unique data and decades of domain expertise accessible to both humans and AI agents, with trust, evidence, and explainability at the core of everything we build. We have partnered with the Apache Software Foundation, the Commonhaus Foundation, the Micronaut Foundation, and other OSS projects such as Spring, Quarkus, Kotlin, JUnit, AndroidX, and many more to bring the values of Develocity also to the OSS Community. Our Values Seek to Understand: Everything starts with listening and understanding, and we strive to understand different viewpoints, problems, and motivations. Before we take action, we ensure we truly grasp the challenges, perspectives, and goals. Know the Why : We approach our work with a clear sense of purpose, ensuring every step is deliberate and focused. We take meaningful action with urgency, but never at the expense of thoughtful consideration. Innovate & Iterate : We embrace challenges and are not afraid to try new things, even if they might fail. With deep understanding and a clear purpose, we can develop creative and bold solutions to tackle challenges. Own the Outcome: We are empowered to take initiative and we maintain transparency in our work and its outcomes. When we execute, we take responsibility for our decisions, measure the success of our innovations, and learn from the results. Who You Are We're building a new SRE team and looking for founding members to help shape how we operate. You'll be responsible for the reliability, performance, and availability of Develocity instances serving paying customers, open-source projects, and public-facing services, plus supporting infrastructure like artifact registries. You'll work on our internally-built Cloud Application Platform, Kubernetes on AWS, and develop deep expertise in it. When incidents happen, you'll troubleshoot issues across the stack, from application to infrastructure. You'll collaborate with the Cloud Platform team to improve the tooling you depend on, and with engineering teams to build reliability into how we ship software. If you like automating things and hate doing the same task twice, you'll fit in well. You'll be part of a distributed, remote-first team that values asynchronous communication and written documentation. Strong self-direction and clear communication across time zones are essential. Responsibilities

* Operate and maintain all Develocity instances and supporting services.

* Participate in a follow-the-sun on-call rotation, owning incident response and troubleshooting issues across the stack.

* Drive automation across application deployment, upgrades, monitoring, self-healing, and recovery.

* Build and maintain observability for all managed services (logging, metrics, tracing, and alerting).

* Work with engineering teams to build reliability into features from the start.

* Run incident response and retrospectives, and make sure we learn from them.

* Own disaster recovery, backups, and business continuity.

* Communicate with customers during incidents and maintenance windows.

* Optimize performance, resource usage, and costs.

* Help evolve our SaaS operations as we grow.

Minimum qualifications

* 5+ years in SRE, DevOps, or equivalent role operating production services at scale.

* Strong Kubernetes experience in production environments.

* Cloud infrastructure expertise, preferably AWS (EKS, RDS, S3, EC2).

* Proficiency with observability tools (Prometheus, Grafana) and Infrastructure as Code (Terraform).

* Track record of incident management and response.

* Knowledge of SRE best practices (SLAs, SLOs).

* Scripting proficiency (Python, Bash) for automation.

* Experience with 24/7 on-call rotations.

* Strong written and verbal English communication.

Preferred qualifications

* Experience operating SaaS platforms at scale.

* Familiarity with Develocity.

* JVM language experience (Java, Kotlin).

* Disaster recovery planning and execution experience.

* Customer-facing incident communication skills.

* Experience establishing SRE practices in new or growing teams.

What We Offer

* A ground-floor role in a new SRE team-you'll shape how we do things, not inherit someone else's decisions.

* Real ownership of production systems used by engineers at companies you've heard of.

* Direct interaction with customers when things go wrong (and when they go right).

* A culture that values automation over heroics.

* In-person meetings, such as our annual company offsite and team meetings.

* Work from home in a remote-first environment.

* Competitive salaries and equity grants.

Location

* Remote from anywhere in Europe in the GMT timezone.

* While our team works remotely and is spread across the globe, we deeply value daily interactions and collaboration.

Originally posted on Himalayas

职位正文保留公司发布时的原文,因为投递的时候用的也是同一种语言。

该职位发布于 Himalayas. 原始招聘信息 (Himalayas)

关于这个职位的常见问题

我可以在自己居住的地方申请「Senior Site Reliability Engineer」这个职位吗?

可以。对于这个职位,Gradle 接受来自欧洲任何国家的候选人,所以你不需要其他国家的工作许可。这条招聘信息通过了自动核查:如果雇主要求工作许可、签证担保,或者必须居住在某个特定国家,它就不会出现在本站。

标明的报酬是多少?

对于这个职位,Gradle 没有公布报酬。多数远程招聘信息不给出具体数字,这件事会在面试时谈定。

怎么申请?

你通过发布在 Himalayas 上的原始招聘信息,直接向雇主提交申请。Donator 不代收申请,不收取佣金,也不保存简历。

这是什么类型的职位?

这是「IT 与编程」类别中的一个完全远程职位。混合办公的招聘信息,以及任何需要到办公室的职位,本站都不会发布。

适合这个职位的免费证书

这个领域里分量最重的几张免费证书。每一张都在发证方自己的页面上核实过。

全部免费证书:「IT and cloud」

相似职位

远程职位:Senior Site Reliability Engineer | Donator