PCI Pal
Site Reliability Engineer

How your CV stacks up
Upload your CV to see how well it fits this job role
?%
Welcome to PCI Pal
PCI Pal is a leading provider of SaaS solutions that empower companies to take payments securely, adhere to strict industry governance, and remove their business from the significant risks posed by non-compliance and data loss. We are integrated and resold by some of the world's leading business communications vendors, as well as major payment service providers.
We are currently looking for a Site Reliability Engineer to join our UK team.
Location: This role will require monthly travel to our London and Ipswich office
The opportunity
You'll join PCI Pal's DevOps team as the go-to authority on observability and reliability, owning our DataDog implementation and extending monitoring, alerting and anomaly detection across infrastructure, networking and telephony. Working closely with the senior DevOps engineer and the wider team, you'll drive down manual toil, improve resiliency, and own capacity planning, availability targets and disaster recovery readiness as load grows. You're joining an existing team and will be relied on to raise the reliability and observability bar for the whole group.
You will be responsible for
- Own the observability framework and standards across infrastructure, networking and telephony, defining what good monitoring and alerting look like, and partnering with development teams to implement against these standards, including a practical monitoring scorecard, reviewing and signing off on their coverage
- Control observability costs - right-sizing usage (hosts, custom metrics, log ingestion/retention, APM) and making deliberate cost-vs-coverage tradeoffs
- Work with teams to reduce alert noise while making sure genuine incidents are still surfaced immediately
- Act as the technical authority on monitoring maturity across the engineering function
- Mentor and upskill the wider DevOps team on reliability and observability practices, working alongside the senior DevOps engineer
- Identify and automate away manual toil, including incident response workflows, compliance evidence collection and release processes
- Drive resiliency improvements across the infrastructure estate - redundancy, failover and capacity planning
- Lead on incident response and resolution, contributing to postmortem practices that turn failures into lasting fixes
- Work with development teams to embed resiliency practices into release processes, e.g. progressive rollouts, canary deployments and fast rollback, so changes can ship quickly without risking availability
Reasons to use Rodeo
I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?
Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.
Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.
Start with a chat, not a search bar
Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.
Graduate Consultant — 2026 Scheme
Why you're a good match
StrongYour economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.
See breakdownIt searches the market for you
Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.
Why you're a good match
You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.
Experience fit
Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.
Only hits
No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.
Job requirements
- Extensive, hands-on experience with DataDog (or equivalent tools such as New Relic, Grafana/Prometheus, Dynatrace)
- Experience managing and controlling observability tooling costs at scale
- Solid experience monitoring networking infrastructure (latency, packet loss, device health, anomaly detection)
- Experience with capacity planning, demand forecasting, and redundancy/failover design
- Strong background in incident response practices and building automation around them
- Experience defining SLIs/SLOs and error-budget-based alerting
- Solid software engineering skills, writing maintainable, testable, version-controlled automation and tooling
- Experience implementing progressive delivery practices (canary deployments, staged rollouts, automated rollback)
- Confidence engaging with and influencing development teams
- Experience of PCI compliance or other similar compliance frameworks
- Experience with automated build systems (e.g. Jenkins), work management systems (e.g. Jira), and source control


Get help with your application
Your very own career expert that helps elevate your application to the next level.
Nice to haves
- Experience with telephony/voice infrastructure monitoring
- Experience mentoring or upskilling other engineers
Talk to us
If you have any questions or want to find out more, we’d love to hear from you.
Please contact the Recruitment Team recruitment@pcipal.com
“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”
Jessica, London
Location