Rodeo
Get started

bet365

Site Reliability Engineer

Stoke-on-Trent
Posted about 16 hours ago
Sign up to applySee more jobs like this

How your CV stacks up

1Upload CV
2Analyse CV
3Improve CV

Upload your CV to see how well it fits this job role

?%

Company Description

We’re one of the world’s leading online gambling companies, revolutionising the industry since 2000. Founded by Denise Coates CBE, we now employ over 10,000 people and serve over 120 million customers in 26 languages.

We empower our employees to push boundaries and explore new ideas, cultivating a culture that celebrates and rewards creativity. This offers employees a wealth of growth opportunities, giving them the opportunity to make a real impact in the world of online gambling. As a forward-thinking company, we’re breaking new ground in software innovation too, redefining what’s possible for our global worldwide.

Our focus on In-Play betting has solidified our market-leading position, featuring more than 1.38 million In-Play sporting events a year. With over 750 concurrent sporting fixtures at peak and more live sports streamed than anyone else in Europe (750,000), we handle over 6 million HTTP requests daily and process more than 1.5 million bets per hour at peak.

Job Description

As a Site Reliability Engineer, you will enhance system reliability, observability, and performance through a strong engineering approach and assist with incident resolution and best practices.

You will have strong software engineering skills, approaching system reliability and observability as a software problem — protecting, providing for, and progressing the performance and availability of our critical systems.

Using your engineering expertise, you will implement solutions that enhance reliability, including service instrumentation with OpenTelemetry and improved logging practices.

You will leverage AI tools and LLM platforms in your daily work to reduce toil, drive autonomous operations, and optimise system health, while engineering automation and tooling for effective service management.

Reasons to use Rodeo

I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?

Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.

Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.

Start with a chat, not a search bar

Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.

P

Graduate Consultant — 2026 Scheme

PwC·London, UK
£35,000/yr

Why you're a good match

Strong

Your economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.

See breakdown
Save jobNot relevant
View details

It searches the market for you

Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.

Why you're a good match

You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.

See breakdown
Strong

Experience fit

Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.

See breakdown
Strong

Only hits

No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.

Collaboration is key, working across multiple functions to embed reliability and observability best practices throughout the software development life cycle. Your contributions will ensure our systems meet user demands and foster a culture of continuous improvement.

This role is eligible for inclusion in the Company's hybrid working from home policy.

Qualifications

  • Excellent knowledge of programming languages including Python, Golang, and JavaScript.
  • Knowledge and experience of modern software development techniques and lifecycles.
  • Excellent knowledge of Site Reliability Engineering (SRE) principles, including the creation and management of effective Service Level Indicators (SLI's) and Service Level Objectives (SLO's) for reliability and customer satisfaction.
  • Knowledge of contemporary observability tools, techniques, and best practice including Splunk, New Relic, Grafana, and PagerDuty.
  • Proficiency in shell scripting for automation and system management tasks.
  • Experience with Infrastructure as Code (IaC), automation, and orchestration tools such as Ansible and Terraform.
  • Prior experience working in a large-scale, 24/7 enterprise where system uptime and stability is of paramount importance to the business.
  • An AI-native engineering approach, with hands-on experience using LLM platforms and coding assistants to improve productivity and quality, and the ability to integrate AI-driven telemetry for advanced observability, predictive insights, and root-cause analysis.

Additional Information

  • Developing and maintaining tools that facilitate effective management of our systems, ensuring they are operationally efficient and resilient.
  • Working with automation and orchestration platforms to automate manual activity and reduce toil.
  • Writing and contributing to code that enhances the reliability and observability of services, including telemetry, operational APIs, and tooling.
  • Building sophisticated dashboards using a range of telemetry data and dashboarding technologies like Grafana, Splunk, and New Relic.
  • Actively participating in live incident resolution and post-mortem analysis, providing effective remediation strategies to improve overall system health and prevent future issues.
  • Driving initiatives to enhance system reliability and observability, contributing to a culture of continuous improvement.
  • Maintaining and administering existing monitoring and analytic toolsets.
  • Mentoring colleagues in the use of new technologies or practices.
  • Working with IT Operations to provide and support the use of critical tooling that will enable increasing levels of value to the Business.

Get help with your application

Your very own career expert that helps elevate your application to the next level.

Get help applying for this job

By applying to us, you are agreeing to share your Personal Data in accordance with our Recruitment Privacy Notice - https://www.bet365careers.com/privacy-policy

At bet365, we're committed to creating an environment where everyone feels welcome, respected, and valued. Where all individuals can grow and develop, regardless of their background. We're Never Ordinary, and we're always striving to be better. If you need any adjustments or accommodations to the recruitment process, at either application or interview, please don’t hesitate to reach out.

Workplace Type: Hybrid
Department: Platform Engineering
Full Time/Part Time: Full Time
Shift Pattern: Days
2nd Office Location: UK - Manchester
Job Type: Standard

Trusted by 25,000+ job seekers

“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”

Jessica, London

Get help applying for this job

Skills

Python
Golang
JavaScript
Site Reliability Engineering
Observability
Splunk
New Relic
Grafana
PagerDuty
Shell Scripting
Ansible
Terraform
Infrastructure as Code
OpenTelemetry
LLM Platforms
Automation

Location

Stoke-on-Trent, England, United Kingdom

Sign up to applySee more jobs like this