Rodeo
Get started

Jobgether

AI Researcher — Distillation

UK
Posted about 14 hours ago
Sign up to applySee more jobs like this
Get notified of more jobs like this · No spam, ever

How your CV stacks up

1Upload CV
2Analyse CV
3Improve CV

Upload your CV to see how well it fits this job role

?%

AI Researcher — Distillation

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI Researcher — Distillation based in the United Kingdom.

Join a highly technical research environment focused on advancing the efficiency and performance of modern AI models.

You will research techniques that transform large, resource-intensive models into smaller, faster, and more deployable systems without sacrificing quality. The role combines fundamental machine learning research with hands-on experimentation and real-world engineering challenges.

You will have the opportunity to explore model distillation across large language models, long-context architectures, and inference-constrained environments. Your work will move from research ideas and rigorous experiments into production systems with measurable impact. You will collaborate closely with engineers while contributing to publications, technical research, and potentially open-source projects.

This opportunity is particularly suited to researchers who want meaningful ownership of their work and the ability to see their ideas deployed in practice.

Accountabilities:

  • Design, implement, and evaluate advanced model distillation techniques, including teacher-student training, self-distillation, layer-wise distillation, and representation matching.
  • Investigate the tradeoffs between model size, latency, memory consumption, throughput, and accuracy.
  • Develop novel approaches to distilling large language models, long-context or specialized architectures, and models designed for inference-constrained environments.
  • Conduct large-scale experiments, ablation studies, and rigorous analysis to validate research hypotheses and identify meaningful improvements.
  • Translate research findings into practical implementations and collaborate closely with engineering teams to productionize successful approaches.
  • Prepare and submit research papers to leading machine learning conferences and venues such as NeurIPS, ICML, ICLR, and COLM.
  • Contribute to internal research documentation, technical articles, and open-source machine learning projects where appropriate.
  • Clearly communicate research objectives, methodologies, results, tradeoffs, and limitations to technical stakeholders.

Reasons to use Rodeo

I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?

Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.

Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.

Start with a chat, not a search bar

Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.

P

Graduate Consultant — 2026 Scheme

PwC·London, UK
£35,000/yr

Why you're a good match

Strong

Your economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.

See breakdown
Save jobNot relevant
View details

It searches the market for you

Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.

Why you're a good match

You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.

See breakdown
Strong

Experience fit

Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.

See breakdown
Strong

Only hits

No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.

Requirements:

  • Strong academic or professional background in machine learning research, with a solid understanding of deep learning fundamentals.
  • Hands-on experience with model distillation or closely related areas such as model compression, pruning, quantization, or representation learning.
  • Demonstrated publication experience through conference or journal papers, workshop publications, or arXiv preprints.
  • Strong understanding of optimization, training dynamics, generalization, and modern deep learning methodologies.
  • Fluency in PyTorch or an equivalent deep learning framework, with experience conducting research-grade experimentation.
  • Ability to design rigorous experiments, interpret results, and critically evaluate research approaches.
  • Strong written and verbal communication skills, with the ability to explain complex research ideas and findings clearly.
  • Experience with large language model distillation is highly valued.
  • Background in efficiency-focused research involving latency, memory, throughput, or related deployment constraints is advantageous.
  • Experience with long-context models or non-Transformer architectures is a plus.
  • Open-source contributions to machine learning, research tooling, or related projects are beneficial.
  • Prior startup or applied research experience is welcome.
  • PhD, postdoctoral, academic research, or industry research experience in machine learning or a related field is particularly relevant, though equivalent research backgrounds may also be considered.

Get help with your application

Your very own career expert that helps elevate your application to the next level.

Get help applying for this job

Benefits:

  • Significant ownership and influence over research direction within a Series A-stage environment.
  • Strong support for publishing research and pursuing open research initiatives.
  • Close feedback loop between research experimentation and real-world production deployment.
  • Access to meaningful compute resources and production-scale machine learning problems.
  • Opportunity to work on cutting-edge model efficiency and distillation challenges.
  • Collaboration within a small, highly technical team with deep expertise across machine learning and systems.
  • Opportunity to see research progress from papers and experimental code through to deployed AI systems.
  • Exposure to large language models, efficient inference, long-context architectures, and other emerging AI technologies.

How Jobgether Works:

We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Why Apply Through Jobgether?

Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

#LI-CL1

Trusted by 25,000+ job seekers

“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”

Jessica, London

Get help applying for this job

Skills

Model distillation
Deep learning
PyTorch
Machine learning research
Model compression
Pruning
Quantization
Representation learning
Large language models
Optimization
Training dynamics
Generalization
Ablation studies
Long-context architectures
Inference-constrained environments

Location

United Kingdom

Sign up to applySee more jobs like this