Hamilton Barnes 🌳
Site Reliability Engineer

How your CV stacks up
Upload your CV to see how well it fits this job role
?%
Job Title: Site Reliability Engineer (GPU Infrastructure)
Location: Remote (UK Wide)
Salary: Up to £120,000 depending on experience
A high growth GPU cloud provider is hiring a Site Reliability Engineer to keep large scale production GPU infrastructure fast, stable and observable. This isn't a sandbox or legacy estate, you'd be working directly on live infrastructure supporting demanding AI workloads, sitting where reliability, networking and platform engineering meet. The role is fully remote from anywhere in the UK, within a small and technically strong engineering led team.
What's in it for you?
- Real production scope from day one on live GPU infrastructure, not a ticket queue or sandbox environment
- Rare breadth across reliability, high performance networking and platform engineering, with room to grow into whichever area interests you most
- Fully remote working from anywhere in the UK
- A team focused on building durable systems and fixing root causes through automation, rather than repeating the same manual fixes every week
Reasons to use Rodeo
I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?
Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.
Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.
Start with a chat, not a search bar
Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.
Graduate Consultant — 2026 Scheme
Why you're a good match
StrongYour economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.
See breakdownIt searches the market for you
Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.
Why you're a good match
You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.
Experience fit
Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.
Only hits
No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.
Responsibilities
- Own the reliability and performance of production GPU infrastructure, including the high performance networking underneath it
- Design, build and maintain observability, monitoring and dashboards, improving signal quality through correlation, enrichment, suppression and deduplication
- Build Python based automation for incident triage, runbook execution and routine operational tasks
- Integrate observability, ITSM and infrastructure APIs to enrich alerts and automate operational workflows
- Deliver internal tools and self service capabilities, including CLI utilities, ChatOps integrations and dashboards
- Turn post incident learnings into better tooling, automation and operational standards


Get help with your application
Your very own career expert that helps elevate your application to the next level.
Skills Required
- Proven track record as an SRE or Production/Infrastructure Engineer with hands on GPU infrastructure in a live production environment
- Strong Python for automation and tooling
- Deep observability and dashboarding experience (Prometheus, Grafana or similar), with alert design and incident management
- UK based and able to work fully remote
Strong Differentiators
- Low latency networking or InfiniBand experience
- Platform engineering background (Kubernetes, infrastructure as code, internal developer platforms)
- HPC infrastructure exposure, such as Slurm managed clusters or research computing environments
- ChatOps and ITSM integration experience (Slack/Teams bots, ServiceNow or Jira Service Management APIs)
Apply now if you want to work on large scale GPU infrastructure with real breadth of scope!
“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”
Jessica, London
Location