Rodeo
ResourcesPartnersSign in

FDM Group

Site Reliability Engineering Lead

London
£90k – £130k/yr
Posted 1 day ago
Sign up to applySee more jobs like this

How your CV stacks up

1Upload CV
2Analyse CV
3Improve CV

Upload your CV to see how well it fits this job role

?%

FDM is a global business and technology consultancy seeking an SRE Lead

FDM is a global business and technology consultancy seeking an SRE Lead to support a major global financial services organisation as it establishes its first formal Site Reliability Engineering function. This is initially a 12 month contract with the potential of going permanent and will be a hybrid role based in Bristol, London or Edinburgh.

This role offers a unique opportunity to build SRE almost from first principles within a large, complex enterprise environment. The successful candidate will lead the transition from largely reactive production support toward a proactive, engineering‑led reliability model, while influencing both legacy platforms and a new, standards‑driven future environment.

You will act as the founding SRE leader, setting the vision, operating model, and priorities for the function, while driving improvements in service stability, resilience, observability, and overall operational maturity. Lead the organisation’s shift from reactive firefighting to data‑driven, preventative reliability practices, and influence platform and application design so that reliability is engineered in from the outset rather than addressed after issues arise.

Responsibilities:

  • Define and embed a scalable SRE operating model aligned to organisational culture and maturity, with clear roles, responsibilities, and ways of working across SRE, platform, infrastructure, and application teams.
  • Build and grow the SRE capability, shaping team structures, backlog priorities, and operating rhythms to improve reliability across both legacy environments and a modern, automated target platform.
  • Establish meaningful service reliability measurement, implementing Critical User Journeys, SLIs, and SLOs to create baselines and provide end‑to‑end service health visibility.
  • Embed SLO‑driven decision making, enabling balanced, data‑led trade‑offs between availability, delivery velocity, and operational risk.
  • Reduce operational toil and increase resilience through automation initiatives, including runbook automation, self‑healing approaches, and improved deployment, patching, as well as recovery strategies.
  • Strengthen incident and problem management, enhancing major incident response, coordination, and communication, and driving high‑quality root cause analysis to prevent recurrence.
  • Define and implement pragmatic observability standards across logging, metrics, tracing, alerting, and dashboards to reduce noise and improve signal quality.
  • Lead stakeholder engagement and organisational change, influencing senior leaders to adopt SRE principles and acting as a trusted authority on reliability and operational excellence.

Reasons to use Rodeo

I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?

Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.

Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.

Start with a chat, not a search bar

Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.

P

Graduate Consultant — 2026 Scheme

PwC·London, UK
£35,000/yr

Why you're a good match

Strong

Your economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.

See breakdown
Save jobNot relevant
View details

It searches the market for you

Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.

Why you're a good match

You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.

See breakdown
Strong

Experience fit

Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.

See breakdown
Strong

Only hits

No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.

About You

Requirements

  • Significant hands‑on experience in Site Reliability Engineering or reliability‑focused production engineering roles.
  • Proven success establishing SRE practices in a large, complex enterprise environment.
  • Experience working in environments with legacy platforms on‑prem or private cloud infrastructure.
  • High operational and regulatory expectations.
  • Strong engineering background with a practical approach to automation and problem solving.
  • Deep understanding of incident management, problem management and operational resilience.
  • Ability to design simple, scalable operating models.
  • Calm, confident leadership during high‑pressure production incidents with clear, credible communication at senior stakeholder level.
  • Pragmatic, outcome‑focused engineering mindset.
  • Foundational leadership in ambiguous, evolving environments.

Get help with your application

Your very own career expert that helps elevate your application to the next level.

Get help applying for this job

Desirable Experience:

  • Experience building SRE capability where maturity was low or undefined.
  • Exposure to large‑scale technology transformation programmes.
  • Experience standardising observability across fragmented toolsets.
  • Familiarity with global, multi‑application environments supporting critical business services.
  • Experience mentoring teams or building communities of practice around reliability.

About Us

FDM is an award-winning global leader in tech and business talent solutions, backed by more than 35 years of industry experience. We have centres across Europe, North America, and Asia-Pacific, and a global workforce of over 2500 employees. FDM has shown exponential growth throughout the years, firmly establishing itself as an award-winning employer, currently listed on the FTSE4Good Index and as a 2026 Financial Times UK ‘Best Employer’.

Diversity and Inclusion

FDM Group is an equal opportunity employer, and all qualified applicants will receive consideration for employment without regard to race, colour, religion, sex, sexual orientation, national origin, age, disability, veteran status or any other status protected by federal, provincial or local laws.

Why join us

  • Career coaching, mentoring and access to upskilling throughout your entire FDM career.
  • Assignments with global companies and opportunities to work abroad.
  • Opportunity to re-skill and up-skill into new areas, develop non-linear career paths and build a skillset within your field.
  • Annual leave and workplace pension.
Trusted by 25,000+ job seekers

“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”

Jessica, London

Get help applying for this job

Skills

Site Reliability Engineering
Incident Management
Problem Management
Observability
Automation
SLIs
SLOs
Critical User Journeys
Operational Resilience
Stakeholder Engagement
Root Cause Analysis
Platform Engineering
Infrastructure Management
Service Reliability
Leadership

Location

London, England, United Kingdom

Sign up to applySee more jobs like this