Rodeo
ResourcesPartnersSign in

Sadler Recruitment

Junior Site Reliability Engineer

Cardiff
£35k – £40k/yr
Posted about 20 hours ago
Sign up to applySee more jobs like this

How your CV stacks up

1Upload CV
2Analyse CV
3Improve CV

Upload your CV to see how well it fits this job role

?%

Role: Junior Site Reliability Engineer
Location: Cardiff, South Wales
Office requirements: Remote (central Cardiff office available 2 days a week)
Salary: £35,000 - £40,000 per annum

About the Role

This company provides managed AI operations for technology businesses. The service is not a traditional managed service model of cloud support and resets. The company operates, secures, and governs the cloud, observability, and AI runtime layer behind mission-critical software. The client's team owns the product, application, and model behaviour; we own the operating layer that keeps it reliable, secure, cost-controlled, and evidence-ready.

Clients range from start-ups and scale-ups building fast without formal engineering controls, through to regulated, mission-critical platforms in fintech, healthtech, and insurance, where the operating model and evidence trail matter as much as uptime. You'll work across traditional cloud platforms like Azure and AWS, as well as newer, AI-native application stacks such as Lovable, and the modern cloud platforms that often sit behind them, like Vercel and Supabase.

The company is a Datadog Advanced Partner (UK) and holds the accolade of being the world's first accredited MSP powered by Datadog. Datadog is a Nasdaq-listed observability platform. The company's technical focus is on observability and LLM observability specifically.

This is a chance to grow with a Cardiff business on the front line of AI-era cloud operations, with bleeding-edge tech, AI used responsibly, a team worth learning from, and the flexibility to do your best work. Other awards include the Sir Michael Moritz Tech Start-up Award.

Key Responsibilities

Runtime Assurance Delivery

  • Support customers across a mix of traditional and modern cloud platforms, Azure and AWS, alongside newer AI-native application stacks and the cloud platforms behind them, such as Lovable, Vercel, and Supabase.
  • Work through an ongoing ticket backlog per customer, sized to allow both reactive support and proactive improvement work.
  • When a customer's backlog is empty, proactively review their environment for improvements rather than waiting for tickets to be raised.

Reasons to use Rodeo

I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?

Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.

Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.

Start with a chat, not a search bar

Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.

P

Graduate Consultant — 2026 Scheme

PwC·London, UK
£35,000/yr

Why you're a good match

Strong

Your economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.

See breakdown
Save jobNot relevant
View details

It searches the market for you

Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.

Why you're a good match

You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.

See breakdown
Strong

Experience fit

Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.

See breakdown
Strong

Only hits

No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.

Observability & Datadog

  • Receive training on client environments and on the Datadog platform.
  • Identify improvements independently, either by raising them to the team or by implementing them directly.
  • Tickets may also be generated directly by Datadog alerts, covering both backlog items and live incidents.

Build & Improvement Projects

  • Deliver end-to-end build-and-run projects for managed service customers, from scoping through to live operation.
  • Implement cloud cost optimisation strategies, working through the technical dependencies that can sometimes block them.
  • Manage monitoring and observability infrastructure as estates grow, keeping platforms performant and well-maintained.
  • Build and deploy new infrastructure in Azure, including infrastructure-as-code, Datadog implementation, and APM setup, covering the full lifecycle from onboarding through go-live stabilisation.

AWS Platform Work

  • Work hands-on with containers, EC2, and RDS instances across a range of customer environments, from day-to-day configuration and troubleshooting through to cost and performance optimisation.
  • Also extends into AWS Bedrock for customers building AI-native products.

Internal Tooling & Process

  • Contribute to a growing suite of internal AI tools built to improve engineering efficiency, reduce context-switching across platforms, and modernise how we run cloud operations day to day.

A Day in the Life

  • Review ticket queues across the managed service accounts.
  • Project time is spent on a mix of Azure, AWS, and Datadog implementation work, alongside newer platforms like Lovable, Vercel, and Supabase.
  • Work includes cost optimisation, infrastructure builds, and Datadog implementation and APM setup on live applications.
  • On AWS, this includes container work, EC2, and RDS configuration, and Bedrock. Python is used throughout.
  • Depending on the customer, this work may also involve GitHub Actions, Azure DevOps pipelines, or PowerShell scripting. Time is also allocated to internal work covering internal AI tools and process improvements.
  • The role involves working across managed service delivery, project builds, observability, and internal tooling within the same week, rather than working on a single area on an ongoing basis.

Get help with your application

Your very own career expert that helps elevate your application to the next level.

Get help applying for this job

Experience Required

  • Solid Linux fundamentals: CLI, networking, process management
  • Working knowledge of at least one cloud platform (AWS or Azure)
  • Comfort with scripting: Bash, Python, or similar
  • Understanding of core observability concepts: metrics, logs, traces
  • Clear written and verbal communication: you'll work with customers

Nice to have

  • Hands-on Datadog experience (any tier)
  • Terraform or other IaC tooling
  • Kubernetes or containerised workload exposure
  • Experience in a managed service provider or multi-customer environment
  • Familiarity with ISO 27001 or similar compliance frameworks
  • Any cloud certification (AWS, Azure, or Datadog)

What's on Offer

  • Join a growing Cardiff business entering its scaling phase, with real ownership, visibility, and a chance to shape how a fast-moving modern MSP operates AI-era cloud infrastructure.
  • On-call is documented, structured, and remunerated in addition to salary.
  • Training and tooling costs are paid for by the company.
  • Attendance at industry events and courses is encouraged and funded as part of ongoing development.
  • Employees are expected to identify and pursue improvements independently, without pre-defined limits on scope.

Due to expected high demand, we'll be reviewing CVs as they come in and may close this role early once we've received sufficient applications.

For immediate consideration, please send your CV in today.

Trusted by 25,000+ job seekers

“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”

Jessica, London

Get help applying for this job

Skills

Linux
Cloud Platforms
AWS
Azure
Python
Bash
Observability
Datadog
Infrastructure as Code
Terraform
Kubernetes
Containers
Networking
Scripting
DevOps
Cloud Cost Optimization

Location

Cardiff, Wales, United Kingdom

Sign up to applySee more jobs like this