Rodeo
Get started

Grafana Labs

Staff Software Engineer - Databases SRE | UK | Remote

United Kingdom (Remote)
£103.9k – £124.8k/yr
Posted 2 days ago
Sign up to applySee more jobs like this
Get notified of more jobs like this · No spam, ever

How your CV stacks up

1Upload CV
2Analyse CV
3Improve CV

Upload your CV to see how well it fits this job role

?%

Staff Software Engineer - SRE (Remote – UK, Sweden, Spain, or Germany)

About the Company

Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With our actually useful AI, organizations can see, understand, and act on their disparate data to move at the speed of their ambitions.

Today, more than 35 million users and 7,000+ customers—including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce—trust Grafana Labs for reliability, incident resolution, and optimized telemetry. We are a 100% remote company with 1,600+ team members across 40+ countries and backed by leading investors.

About the Role

We are scaling fast and staying true to what makes us different: an open-source legacy, a global collaborative culture, and a passion for meaningful work. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do.

This is a remote opportunity targeting candidates from the UK, Sweden, Spain, or Germany.

We are seeking a Staff Software Engineer - SRE to help support Grafana Cloud’s highest-value customers by ensuring the reliability of our Mimir, Loki, Tempo, and Pyroscope databases. These databases are provided as a SaaS product across AWS, GCP, and Azure globally.

As an SRE, you will:

  • Partner closely with product engineering squads (embedded model)
  • Own production reliability for high-SLA and complex customer environments
  • Design and implement automation to scale our reliability practices
  • Ensure customers meet SLO targets
  • Define and evolve per-tenant SLOs and reliability models
  • Proactively reduce SLO burn to prevent repeat incidents
  • Serve as a primary escalation and on-call point for incidents
  • Lead incident response and Post Incident Reviews (PIRs)
  • Influence feature design to improve production scalability and operability
  • Believe automation to eliminate toil where needed
  • Improve observability and alert quality

Reasons to use Rodeo

I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?

Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.

Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.

Start with a chat, not a search bar

Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.

P

Graduate Consultant — 2026 Scheme

PwC·London, UK
£35,000/yr

Why you're a good match

Strong

Your economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.

See breakdown
Save jobNot relevant
View details

It searches the market for you

Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.

Why you're a good match

You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.

See breakdown
Strong

Experience fit

Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.

See breakdown
Strong

Only hits

No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.


Key Responsibilities (Daily Work)

  • Conduct regular 1:1s with your manager and team
  • Review, create, and proactively optimize SLOs
  • Implement system improvements: monitoring enhancements, automation, self-healing, and auto-scaling
  • Enhance customer observability in their environments
  • Design reliable and scalable solutions to meet growing demands
  • Advocate for fault-tolerant architecture throughout the service lifecycle
  • Collaborate with Engineering Leaders to shape product strategy, roadmaps, and technical decisions
  • Review PRs and Design Docs while mentoring other engineers
  • Teach Site Reliability Engineering principles to help feature development early
  • Participate in incident response, investigations, PIRs, and customer communications

Requirements & Preferences

Experience and Technical Skills

✅ 8+ years in engineering, with 4+ years in SRE/CRE/production engineering (preference for formal customer reliability engineering experience). ✅ Strong Kubernetes expertise in AWS, GCP, or Azure + familiarity with Infrastructure-as-Code tooling: Help • Terraform • Jsonnet • Helm ✅ Technical leadership experience, including:

  • Leading projects
  • Mentoring engineers ✅ Operating multi-tenant systems in production ✅ Deep experience with SLO design + associated reliability monitoring ✅ Programming experience (e.g., Go, Python, Java) ✅ Linux internals, networking fundamentals, cloud storage, and scaling knowledge ✅ Proven troubleshooting and problem-solving skills ✅ Experience with calm, blame-free incident response, including:
  • Following up on actions
  • Writing high-quality PIRs

Get help with your application

Your very own career expert that helps elevate your application to the next level.

Get help applying for this job

✅ Strong reasoning about performance, scaling, and failure modes ✅ Autonomy and self-direction comfort in an engineering-first culture ✅ Agility in partnering with product engineering teams

Soft Skills We Value: ⚡ Intellectual curiosity ⚡ Default-to-transparency culture ⚡ High bias for action ⚡ Kindness (because it matters)


Compensation & Benefits (UK-based Range)

  • Base compensation: £103,958 - £124,750
  • Variable pay: Includes equity, bonus (where applicable), and perks

Note: Compensation adjusts based on location and individual circumstances.

Global Benefits: ✔ Fully remote, global culture (employees worldwide collaborate in a low-visibility, high-trust way). ✔ 30 days annual leave (including 3 Grafana Shutdown Days for full team disconnection). ✔ Open-source roots: We book long-term roadmaps around our open values. ✔ Career growth paths: Clear progression opportunities.


Why We’re Thrilled to Have You

⭐ Remote-first culture unboxing global talent and collaboration ⭐ High-growth environment with constant innovation growth ⭐ Transparent communication (with company-wide updates!) ⭐ Autonomy & flexibility to present work, experiment, and scale ⭐ Empowered teams: High trust, low ego team commitments ⭐ In-person onboarding: Integrate quickly and fully with Grafana Shutdown Days.

Meet the Grafana "Grafanistas"—smart, supportive, and passionate about what they do.


Equal Opportunity Employer

Grafana Labs values equality and diversity. We recognize that difference builds a strong organization. Our inclusive culture is key to growth.

Privacy Policy: "Grafana Labs may utilize AI tools in recruitment to help match skills to roles—extension always occurs in human review."

#LI-Remote

Trusted by 25,000+ job seekers

“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”

Jessica, London

Get help applying for this job

Location

United Kingdom

Sign up to applySee more jobs like this