Rodeo
Get started

Verda

Platform Engineer - LLM Inference Infrastructure

Berlin
Posted about 16 hours ago
Sign up to applySee more jobs like this
Get notified of more jobs like this · No spam, ever

How your CV stacks up

1Upload CV
2Analyse CV
3Improve CV

Upload your CV to see how well it fits this job role

?%

At Verda, we're building a full-stack AI cloud, covering everything from data centers and hardware to our own cloud platform that the world's leading AI teams use to do serious AI work.

We strive to make a positive mark on the world through the infrastructure we build and give leading teams a service they can truly depend on. Headquartered in Helsinki, we operate globally with offices in London and San Francisco.

Join Verda while it’s still being built - not once it’s finished.

About the role

We run a multi-tenant LLM inference platform: customers send requests to an OpenAI-compatible gateway, and we handle routing, tenancy, access control, usage metering and billing on top of our own GPU fleet.

You would own backend services and the Kubernetes platform they run on. This is not a role where infrastructure is someone else’s problem, you write the Go service, the Helm chart, the network policy and the runbook, and you are the one who verifies it in production.

We are moving deliberately toward a Kubernetes-native architecture: less imperative tooling and hand-run scripts, more declarative APIs, custom resources and controllers that reconcile state. If you have wanted to build operators and control planes rather than consume them, that is the direction of this role.

Your responsibilities

  • Backend services in Go - the platform API, the customer-facing usage API, the metering and billing pipeline. Small, focused services with real correctness requirements.
  • Kubernetes-native platform work - GitOps deployment, custom resources and controllers, progressive rollout, network policy, secret and certificate management. Moving what is currently scripted into something that reconciles.
  • Multi-tenancy and access control - tenant provisioning, per-project model access, credential handling across environments.
  • Usage metering and billing correctness - a pipeline that turns raw requests into per-tenant token accounting, plus the reconciliation that proves what we bill matches what actually happened. This is money, and it has to be right.
  • Observability - logs, metrics and traces that answer questions during an incident rather than after it.

Reasons to use Rodeo

I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?

Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.

Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.

Start with a chat, not a search bar

Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.

P

Graduate Consultant — 2026 Scheme

PwC·London, UK
£35,000/yr

Why you're a good match

Strong

Your economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.

See breakdown
Save jobNot relevant
View details

It searches the market for you

Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.

Why you're a good match

You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.

See breakdown
Strong

Experience fit

Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.

See breakdown
Strong

Only hits

No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.

Your key competencies

  • 4+ years of experience
  • Strong Go. You have shipped and maintained production Go services, and you are comfortable with concurrency, context propagation and error handling that fails loudly instead of silently.
  • Real Kubernetes depth. Not just kubectl apply. You understand the control loop, know why a pod is not ready without guessing, and have written Helm charts, network policies and RBAC that you then had to debug.
  • SQL and relational data modelling. PostgreSQL specifically. You can reason about transactions, indexes and migration safety.
  • You verify your work. You do not report something as working because it deployed and the health check is green. You go and prove it, and you say plainly what you did not test.
  • Clear written English. Design notes, runbooks, incident write-ups. Much of our engineering context lives in writing.

Nice to have

  • Building Kubernetes operators / controllers (controller-runtime, CRDs, kubebuilder, Operator SDK)
  • GitOps at scale - Argo CD or Flux, ApplicationSets, multi-cluster
  • LLM serving internals - vLLM, SGLang, TensorRT-LLM, KServe, or similar; GPU scheduling, batching, KV cache behaviour
  • Distributed messaging (NATS, Kafka) and event-driven pipelines
  • Traefik or Envoy/Istio at the ingress layer
  • Time-series and log stores (VictoriaMetrics/VictoriaLogs, Prometheus, ClickHouse)
  • Frontend competence (React + TypeScript) - our operator console is ours to maintain, and being able to fix it end to end is valuable
  • Billing, metering or payments systems, or anything else where being wrong is expensive
  • Python, for the gateway extension layer

Get help with your application

Your very own career expert that helps elevate your application to the next level.

Get help applying for this job

Why Verda

  • Cash and equity compensation along with various fringe benefits (healthcare, lunch, wellbeing, and more).
  • Profitable operations with rapid, sustained growth.
  • 40+ nationalities, with 6 different ones on the management team.
  • A real chance to make an impact and work alongside world class engineers, researchers, and partners across the global AI ecosystem.

Practicalities

  • Work mode: Based in Helsinki / London or remote in Europe
  • Level: Senior
  • Employment type: Full time and permanent

What's next

We're building fast and this role needs the right person behind it. There's no artificial deadline, but when we find who we're looking for, we move. If this sounds like your next move, apply now.

Please submit your application through our Careers page. We don't accept applications sent by email.

Trusted by 25,000+ job seekers

“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”

Jessica, London

Get help applying for this job

Skills

Go
Kubernetes
PostgreSQL
Helm
GitOps
API development
Observability
Multi-tenancy
Network policy
RBAC
SQL
Distributed systems
Infrastructure as code
Metering
Billing systems

Location

Töölönlahdenkatu 2, 00100 Helsinki, Finland

Sign up to applySee more jobs like this