Rodeo
Get started

Laine

Senior Site Reliability Engineer (SRE) (Dubai relocation required)

London
Posted about 19 hours ago
Sign up to applySee more jobs like this
Get notified of more jobs like this · No spam, ever

How your CV stacks up

1Upload CV
2Analyse CV
3Improve CV

Upload your CV to see how well it fits this job role

?%

About Laine.ai

Laine builds AI-powered workflows for the legal industry, where strict data security, high availability, and deterministic latency are non-negotiable. Our platform processes highly sensitive legal documents and coordinates multi-step LLM operations in real time. We are seeking a Senior Site Reliability Engineer to take ownership of our cloud infrastructure, observability, and deployment pipelines as we scale.

The Role

As our primary Site Reliability Engineer, you will bridge software engineering and infrastructure operations. You will define our reliability roadmap, harden production workloads on Google Cloud Platform, and design resilient pipelines for both traditional API services and latency-sensitive LLM workloads. You will work directly with core engineering to ensure high uptime, rapid incident resolution, and smooth continuous delivery. You will be working from Dubai.

Key Responsibilities

  • Infrastructure as Code & Cloud Architecture: Design, provision, and maintain secure, reproducible cloud infrastructure across GCP (Cloud Run, Cloud SQL / AlloyDB, VPCs,...).
  • Database Reliability & Scalability: Optimize PostgreSQL/AlloyDB clusters, manage connection pooling, tune query performance, implement zero-downtime database migrations with Prisma, and maintain rigorous backup/recovery strategies.
  • LLM Pipeline & Service Reliability: Monitor and optimize uptime, rate limits, and latency profiles for AI model inference (Vertex AI and external LLM APIs), including tracing via LLMOps platforms (e.g., Langfuse).
  • Observability & Alerting: Build centralized telemetry and tracing pipelines using OpenTelemetry, Sentry, Prometheus/Grafana, or GCP Cloud Monitoring. Establish actionable SLOs, SLIs, and alert routing.
  • CI/CD & Developer Enablement: Maintain fast, reliable automated build, test, and release pipelines using Cloud Build and containerized workflows.
  • Security & Compliance Posture: Enforce least-privilege IAM policies, manage secrets securely, ensure zero public exposure of sensitive databases, and maintain compliance standards required for enterprise legal technology (SOC 2, ISO 27001).
  • Incident Management: Lead post-mortems, identify root causes, implement preventative safeguards, and define the team's incident escalation processes.

Reasons to use Rodeo

I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?

Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.

Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.

Start with a chat, not a search bar

Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.

P

Graduate Consultant — 2026 Scheme

PwC·London, UK
£35,000/yr

Why you're a good match

Strong

Your economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.

See breakdown
Save jobNot relevant
View details

It searches the market for you

Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.

Why you're a good match

You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.

See breakdown
Strong

Experience fit

Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.

See breakdown
Strong

Only hits

No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.

What We Are Looking For

  • 4+ years of professional SRE or DevOps experience, ideally within a high-growth SaaS or startup environment.
  • Deep GCP expertise: Hands-on experience configuring and debugging containerized workloads (Cloud Run), networking, IAM, and security perimeters.
  • Relational Database Administration: Strong production experience tuning, monitoring, and scaling PostgreSQL or AlloyDB under heavy read/write loads.
  • Software Engineering Fundamentals: Fluency with modern backend architectures, particularly TypeScript/Node.js (NestJS ecosystem) and Docker-based containerization.
  • Infrastructure as Code: Proven track record managing production state and multi-environment setups.
  • Production Observability: Experience architecting monitoring stacks that provide end-to-end visibility into distributed services and async job queues.
  • Security-First Mindset: Practical experience locking down data layers, encryption at rest/in transit, and audit logging for sensitive enterprise customer data.
  • Moving to Dubai: You must be happy relocating to Dubai, United Arab Emirates.

Get help with your application

Your very own career expert that helps elevate your application to the next level.

Get help applying for this job

On-Call & 24/7 Production Support

Because Laine provides mission-critical infrastructure to enterprise legal teams globally, maintaining high availability and rapid incident response is critical.

  • 24/7 Rotating Schedule: You will participate in a scheduled 24/7 on-call rotation shared across the engineering team, serving as a designated primary or secondary responder for production incidents.
  • Response SLAs: Respond promptly to high-priority automated alerts (P1/P2) during your on-call shift, triaging critical outages, service degradations, and data pipeline failures.
  • Escalation & Tooling: Leverage modern incident response tools (e.g., PagerDuty, Opsgenie, GCP Cloud Monitoring alerts) to triage issues and communicate incident status transparently.
  • Sustainable Operations: We treat every page as an opportunity to fix the underlying system. You will lead blameless post-mortems, automate failure remediation, and refine alert thresholds to prevent repeat incidents and protect on-call sustainability.

Nice to Have

  • Experience monitoring and optimizing LLM applications, token streaming latency, and vector search operations.
  • Experience implementing or maintaining compliance frameworks (SOC 2 Type II, ISO27001, GDPR, FADP and equivalents).
  • Prior background in legal tech, enterprise SaaS, or data-intensive contextual AI systems.

Tech Stack

  • Cloud & Runtime: Google Cloud Platform, Cloud Run, Docker
  • Data & Storage: AlloyDB / PostgreSQL, Supabase, Redis
  • Backend: TypeScript, NestJS, Prisma ORM
  • AI & LLM Services: Vertex AI
  • Tooling & CI/CD: Cloud Build, Graphite
  • Observability: Sentry, Langfuse
Trusted by 25,000+ job seekers

“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”

Jessica, London

Get help applying for this job

Location

London, England, United Kingdom

Sign up to applySee more jobs like this