Gradle Technologies
Lead Site Reliability Engineer

How your CV stacks up
Upload your CV to see how well it fits this job role
?%
Who We Are
AI is changing how software gets built. Code production is becoming a commodity. The focus is shifting from writing code to orchestrating, verifying, and governing change – and the toolchain is the new constraint.
We are at the center of this shift. We build Develocity, a toolchain observability and intelligence platform used by some of the world's leading software organizations – Netflix, Airbnb, Spotify, SAP, major global banks, and hundreds more. Develocity helps software teams achieve delivery excellence through deep observability, build and test acceleration, and AI-powered intelligence across the entire toolchain – with current support for Gradle Build Tool, Apache Maven™, sbt, npm, and Python.
We are an AI-native company. AI is not a feature we're bolting on – it's central to how we work, how we think about our product, and where we're heading. We're investing deeply in making Develocity's unique data and decades of domain expertise accessible to both humans and AI agents, with trust, evidence, and explainability at the core of everything we build.
We have partnered with the Apache Software Foundation, the Commonhaus Foundation, the Micronaut Foundation, and other OSS projects such as Spring, Quarkus, Kotlin, JUnit, AndroidX, and many more to bring the values of Develocity also to the OSS Community.
Our Values
- Seek to Understand: Everything starts with listening and understanding; we strive to understand diverse viewpoints, problems, and motivations. Before we take action, we ensure we truly grasp the challenges, perspectives, and goals.
- Know the Why: We approach our work with a clear sense of purpose, ensuring every step is deliberate and focused. We take meaningful action with urgency, but never at the expense of thoughtful consideration.
- Innovate & Iterate: We embrace challenges and are not afraid to try new things, even if they might fail. With a deep understanding and a clear purpose, we can develop creative, bold solutions to tackle challenges.
- Own the Outcome: We are empowered to take initiative, and we maintain transparency in our work and its outcomes. When we execute, we take responsibility for our decisions, measure the success of our innovations, and learn from the results.
Who You Are
We're building a new SRE team and looking for founding members to help shape how we operate. As a Lead SRE, you’ll be a technical and operational leader for reliability across Develocity. You’ll help define our SRE vision, set standards for how we operate production services, and mentor other SREs as the team grows. This is a hands-on role with broad influence across engineering, cloud platform, and customer-facing teams.
Reasons to use Rodeo
I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?
Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.
Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.
Start with a chat, not a search bar
Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.
Graduate Consultant — 2026 Scheme
Why you're a good match
StrongYour economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.
See breakdownIt searches the market for you
Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.
Why you're a good match
You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.
Experience fit
Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.
Only hits
No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.
The SRE team will be responsible for the reliability, performance, and availability of Develocity instances serving paying customers, open-source projects, and public-facing services, plus supporting infrastructure like artifact registries.
You'll work on our internally-built Cloud Application Platform, Kubernetes on AWS, and develop deep expertise in it. When incidents happen, you'll troubleshoot issues across the stack, from application to infrastructure. You'll collaborate with the Cloud Platform team to improve the tooling you depend on, and with engineering teams to build reliability into how we ship software. If you like automating things and hate doing the same task twice, you'll fit in well.
You'll be part of a distributed, remote-first team that values asynchronous communication and written documentation. Strong self-direction and clear communication across time zones are essential.
Responsibilities
- Operate and maintain all Develocity instances and supporting services in production.
- Define and evolve SRE standards, practices, and operating models, including on-call, incident response, postmortems, and SLOs.
- Participate in a follow-the-sun on-call rotation, acting as a technical escalation point for complex or high-severity incidents.
- Lead incident response and blameless retrospectives, ensuring learnings result in measurable reliability improvements.
- Set reliability priorities using risk, customer impact, business goals, SLOs, and error budgets.
- Identify systemic reliability risks and continuously evolve Develocity’s SaaS operations as the platform and customer base grow.
- Lead and influence architectural and design reviews to ensure reliability, scalability, and operability.
- Drive automation across deployment, upgrades, monitoring, self-healing, recovery, and operational workflows.
- Build and maintain comprehensive observability for all managed services, including logging, metrics, tracing, and alerting.
- Own disaster recovery, backups, and business continuity planning and execution.
- Partner with engineering leadership to balance feature delivery with reliability and operational excellence.
- Mentor and coach SREs, supporting technical growth and strong operational practices.
- Help onboard new SREs and contribute to hiring by defining and assessing SRE excellence at Develocity.
- Communicate clearly with customers during incidents and maintenance windows.
- Optimize performance, resource utilization, and operational costs.


Get help with your application
Your very own career expert that helps elevate your application to the next level.
Minimum qualifications
- 7+ years in SRE, DevOps, or an equivalent role operating production services at scale.
- Experience leading reliability initiatives across multiple teams or services.
- Demonstrated ability to influence technical direction without direct authority.
- Experience designing and operating systems with SLOs and error budgets, and exercising strong judgment in balancing reliability, velocity, and cost.
- Strong Kubernetes experience in production environments.
- Cloud infrastructure expertise, preferably AWS (EKS, RDS, S3, EC2).
- Proficiency with observability tools (Prometheus, Grafana) and Infrastructure as Code (Terraform).
- Track record of incident management and response in a 24/7 on-call environment.
- Scripting proficiency (Python, Bash) for automation.
- Strong written and verbal English communication skills.
Preferred qualifications
- Experience as a founding or early SRE establishing practices in a growing SaaS organization.
- Familiarity with Develocity.
- JVM language experience (Java, Kotlin).
- Experience with customer-facing and executive-level incident communications.
What We Offer
- A ground-floor role in a new SRE team - you'll shape how we do things, not inherit someone else's decisions.
- Real ownership of production systems used by engineers at companies you've heard of.
- Direct interaction with customers when things go wrong (and when they go right).
- A culture that values automation over heroics.
- In-person meetings, such as our annual company offsite and team meetings.
- Work from home in a remote-first environment.
- Competitive salaries and equity grants.
Location
Remote from anywhere in Europe (GMT). While our team works remotely and is spread across the globe, we deeply value daily interactions and collaboration.
“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”
Jessica, London
Skills
Location