Rodeo
Get started

MUFG

Vice President, Windows Site Reliability Engineer

London
Posted 1 day ago
Sign up to applySee more jobs like this
Get notified of more jobs like this · No spam, ever

How your CV stacks up

1Upload CV
2Analyse CV
3Improve CV

Upload your CV to see how well it fits this job role

?%

Do you want your voice heard and your actions to count? Discover your opportunity with Mitsubishi UFJ Financial Group (MUFG), one of the world’s leading financial groups. Across the globe, we’re 150,000 colleagues, striving to make a difference for every client, organization, and community we serve. We stand for our values, building long-term relationships, serving society, and fostering shared and sustainable growth for a better world.

With a vision to be the world’s most trusted financial group, it’s part of our culture to put people first, listen to new and diverse ideas and collaborate toward greater innovation, speed and agility. This means investing in talent, technologies, and tools that empower you to own your career.

Join MUFG, where being inspired is expected and making a meaningful impact is rewarded.

Site Reliability Engineering

Site Reliability Engineering are responsible for delivering continuous improvement, automation and self-service offerings to operational teams across Bank EMEA and Securities International.

NUMBER OF DIRECT REPORTS

0

MAIN PURPOSE OF THE ROLE

Responsible for the reliability and efficiency of infrastructure through the delivery of common, repeatable tools and processes that greatly reduce the amount of toil operations must perform.

Member of L3 Engineering team providing subject matter expertise and ultimate escalation.

KEY RESPONSIBILITIES

Primary:

  • Develop software to make infrastructure services self-managing and self-service
  • Deliver continuous service improvement by developing Infrastructure as Code
  • Eliminate manual, repetitive, automatable, tactical tasks that are devoid from value
  • Improve system performance, make effective use of resources, distribute load and reduce latency
  • Identify SLO’s (Service Level Objectives) to meet availability and latency objectives
  • Develop pro-active monitoring solutions that alert on symptoms and not just on outages
  • Perform detailed root cause analysis (RCA’s) on incidents and outages to prevent future
  • Partner with development teams to improve services via rigorous testing and release procedures
  • Identify technical debt and partner with application teams to build remediation plans
  • Develop standard operational procedures and produce effective documentation
  • Analyse workloads and devise suitable cloud migration strategies where appropriate
  • Ensure all project / investment workloads are delivered according to plans and budget defined
  • Liaise with Infrastructure Control and IT Risk teams to satisfy internal and external audit requests
  • Deputise for team lead when required to do so and act-up accordingly
  • Identify cost saving and optimisation opportunities across the group
  • Build strong working relationships across the organisation
  • Adhere to the core values of the bank

Reasons to use Rodeo

I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?

Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.

Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.

Start with a chat, not a search bar

Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.

P

Graduate Consultant — 2026 Scheme

PwC·London, UK
£35,000/yr

Why you're a good match

Strong

Your economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.

See breakdown
Save jobNot relevant
View details

It searches the market for you

Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.

Why you're a good match

You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.

See breakdown
Strong

Experience fit

Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.

See breakdown
Strong

Only hits

No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.

Secondary:

  • Perform daily health and compliance checks for all systems as required
  • Ensure all systems are backed up successfully and any issues are promptly resolved
  • Validate monitoring alerts and batch job failures are detected promptly and satisfactorily resolved
  • Ensure sufficient capacity is available to accommodate drive growth
  • Respond to emails sent to the team distribution list / mailboxes in a timely manner
  • Handle incidents and requests with efficiency and a “customer first” mindset
  • Maintain infrastructure in a highly available, reliable, secure and performant manner
  • General Server / Database / Virtualisation Administration maintenance activities
  • Provide technical support to application support and development teams
  • Provide consultancy to application support and development teams
  • Take part in On-Call & weekend work rotation; triaging and addressing production issues as they arise

SKILLS AND EXPERIENCE

Essential:

  • Exceptional skills in Microsoft Windows Server internals and related technologies
  • Excellent skills in managing and maintaining Active Directory, DHCP, DNS, LDAP and Kerberos
  • Extensive experience in hardware performance monitoring and tuning complex low latency systems.
  • Agile, Site Reliability Engineering (SRE) and DevOps Principles and practices
  • Exceptional knowledge of scripting and programming languages such as PowerShell, Python and C#
  • Fluent in Backup and Recovery processes and procedures
  • Advanced knowledge of Clustering, High-Availability, Replication and Disaster Recovery techniques
  • Ability to tune Network, Storage, Server and Virtualisation layers for optimal performance and reliability
  • Excellent Performance Tuning skills, in-depth knowledge of system internals, performance counters and performance measurement and analysis tools.
  • Ability to interpret and implement CIS security hardening recommendations in a controlled manner
  • Acute awareness of Security and Auditing requirements in a regulated environment
  • “Infrastructure as Code” Principles and practices.
  • “Continuous Integration (CI) and Continuous Development (CD)” Principles and practices
  • Git, Ansible, Terraform and TeamCity
  • Serena Deployment Automation (SDA) and Jenkins

Get help with your application

Your very own career expert that helps elevate your application to the next level.

Get help applying for this job

Highly Desirable:

  • Experience on writing, managing plays/playbooks on AWX / Ansible Tower
  • Advance working knowledge of Kubernetes and Docker container orchestration
  • Microsoft SQL Server, Oracle, Sybase ASE, MongoDB and Snowflake
  • IBM Tivoli / Netcool
  • Nutanix HCI and VMWare ESX
  • Networking Protocols (TCP/IP, DNS, DHCP, VLAN’s)
  • RHEL, Oracle Linux, Oracle Solaris and related technologies
  • Cloud computing - IaaS, PaaS and SaaS offerings across Azure, AWS, GCP and Oracle
  • Knowledge of data security governance and regulations such as GDPR and SOX

Desirable:

  • Dell EMC PowerStore (SAN) and Isilon (NAS)
  • Rubrik, EMC Networker, Data Domain and IBM Tivoli Storage Manager
  • CyberArk
  • Splunk
  • Qualys
  • Cisco Tetration
  • ServiceNow
  • JIRA and Confluence

PERSONAL REQUIREMENTS

  • Excellent communication and interpersonal skills
  • Ability to handle pressure during outages and systematically resolve issues
  • Excellent problem-solving skills
  • Results driven, with a strong sense of accountability
  • A proactive, motivated approach
  • The ability to operate with urgency and prioritise work accordingly
  • A structured and logical approach to work
  • Attention to detail and accuracy
  • Ability to perform well in a pressurised environment
  • Ability to manage constructive conflict effectively
  • The ability to manage large workloads and tight deadlines
  • Able to communicate complex technical concepts to non-technical persons at all levels

We are open to considering flexible working requests in line with organisational requirements.

MUFG is committed to embracing diversity and building an inclusive culture where all employees are valued, respected and their opinions count. We support the principles of equality, diversity and inclusion in recruitment and employment, and oppose all forms of discrimination on the grounds of age, sex, gender, sexual orientation, disability, pregnancy and maternity, race, gender reassignment, religion or belief and marriage or civil partnership.

We make our recruitment decisions in a non-discriminatory manner in accordance with our commitment to identifying the right skills for the right role and our obligations under the law.

Trusted by 25,000+ job seekers

“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”

Jessica, London

Get help applying for this job

Skills

Windows Server
Active Directory
PowerShell
Python
C#
Site Reliability Engineering
DevOps
Infrastructure as Code
Ansible
Terraform
Kubernetes
Docker
Performance Tuning
Disaster Recovery
CI/CD
Git

Location

London, England, United Kingdom

Sign up to applySee more jobs like this