Amelco Limited
Site Reliability Engineer

How your CV stacks up
Upload your CV to see how well it fits this job role
?%
Role: Site Reliability Engineer
Type: Full-time permanent role
Location: Hybrid/ Shoreditch, 3 days per week
About Us
Amelco Ltd are a leading gaming and gambling solution software provider with a strong presence in the USA, UK, and Europe. Through partnerships with global gaming companies, we build cutting-edge technical platforms across sportsbooks, lottery, casino, virtual gaming, and financial trading. Our vision is to shape the future of gaming by transforming operations into intelligent, data-driven solutions that deliver exceptional customer experiences and create sustainable value for all stakeholders. We believe in teamwork, knowledge sharing, and transparency with accountability.
The Role
We’re looking for a hands-on Site Reliability Engineer (SRE) to own the reliability, observability, and cost efficiency of our high-throughput betting platform. You’ll be embedded in our production operations, working directly with development teams to build resilient systems, implement actionable observability, and drive incident response from detection to remediation.
This role is 60% infrastructure automation and 40% application-focused reliability work — you’ll need to dive deep into both our Java Spring Boot services and Kubernetes deployment patterns to build lasting improvements.
What You’ll Work With
Infrastructure & Platform:
- Kubernetes: AWS EKS clusters and on-prem deployments
- GitOps: FluxCD for declarative Kubernetes management across 20+ environments
- Infrastructure as Code: Terraform for AWS resources and Kubernetes configurations
- CI/CD: GitHub Actions pipelines with automated AI code review workflows
Observability Stack:
- Monitoring: Prometheus metrics with custom Java Micrometer instrumentation
- Logging: Loki for distributed log aggregation
- Tracing: Tempo for distributed tracing
- AWS: Cloudwatch
Application Environment:
Reasons to use Rodeo
I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?
Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.
Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.
Start with a chat, not a search bar
Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.
Graduate Consultant — 2026 Scheme
Why you're a good match
StrongYour economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.
See breakdownIt searches the market for you
Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.
Why you're a good match
You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.
Experience fit
Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.
Only hits
No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.
- Core Platform: Java Spring Boot microservices for betting, trading, and customer management
- Event Processing: High-throughput event queues with latency monitoring (Kafka, JMS)
- Data Pipeline: Avro-based S3 buffer systems with fault-tolerant write-ahead logs (StatefulS3Buffer)
- DB: Postgres DB on RDS/Aurora.
Key Responsibilities
Reliability Engineering (40%):
- Partner with development teams to define and manage SLOs/SLIs specific to betting platform services (latency, queue depth, bet processing success rates)
- Enhance observability of our Java Spring Boot services — ensure metrics, logs, and tracing are actionable for detecting and fixing production issues
- Implement chaos engineering experiments targeting our event queue systems and high-availability betting services
- Design and execute pre-deployment readiness checks and post-release validation for risk-critical services
Infrastructure & Platform (40%):
- Own Kubernetes cluster reliability across development, QA, and production environments
- Automate operational processes using Python and Bash scripting within our GitHub Actions ecosystem
- Optimize infrastructure costs through rightsizing EKS workloads, tuning autoscaling policies, and implementing efficient resource utilization patterns
- Strengthen platform guardrails through Flux GitOps policies and Terraform module validation
Incident & Operations (20%):
- Contribute to major incident response for betting platform outages, providing engineering expertise on Java service behavior and infrastructure dependencies
- Design and implement automated remediation patterns for common failure modes (event queue backpressure, connection pool exhaustion, database connectivity)
- Build hourly observability agent rules and thresholds for early-warning detection of production anomalies
- Develop runbooks and automation for common operational tasks across our multi-environment platform


Get help with your application
Your very own career expert that helps elevate your application to the next level.
Required Skills & Experience
- 3+ years in SRE, Platform Engineering, or DevOps roles with hands-on production experience
- Strong Kubernetes operational expertise (AWS EKS, GitOps with Flux, Helm packaging)
- Proven track record building observability for Java Spring Boot applications using Prometheus, Grafana, and Loki
- Infrastructure as Code proficiency with Terraform and GitOps workflows
- Python and Bash scripting skills for automation and tooling development
- Experience designing and operating monitoring for high-throughput event processing systems
- Demonstrated ability to balance infrastructure cost efficiency with betting platform reliability requirements
- Excellent communication skills and ability to work across development, platform, and incident management teams
Nice-to-Have Skills
- Experience with betting/gaming platform architecture and reliability patterns
- Familiarity with Avro-based data pipelines and S3 storage optimization
- AWS Certifications (Solutions Architect, DevOps Engineer, or AWS Certified Kubernetes Administrator)
- Background in chaos engineering for stateful event processing systems
- Knowledge of payment processing and risk calculation system reliability patterns
- Experience with automated incident detection and response workflows
Amelco Benefits
- Pension Scheme: Amelco matches up to 7% contributions of your base salary for all staff. You will automatically be entered at 4%.
- Staff Benefits Scheme: access to the staff benefits portal after successful completion of a 4-month probation period.
- Yearly discretionary bonus scheme and pay reviews
- Opportunity to travel and visit our office in Poland and Hungary
If you are interested, hit the apply button, and please ensure you upload a copy of your CV.
We look forward to hearing from you!
“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”
Jessica, London
Skills
Location