Sundayy
Machine Learning Engineer

How your CV stacks up
Upload your CV to see how well it fits this job role
?%
About The Company
Cantina is an innovative social platform founded by Sean Parker, dedicated to revolutionizing online social interactions through advanced artificial intelligence technology. Our platform features the most sophisticated AI character creator available, enabling users to craft lifelike bots that can engage across voice, video, and text channels. These social AI entities serve as personalized content creators and interactive companions, capable of sharing customized experiences and facilitating seamless group conversations. At Cantina, we are passionate about pushing the boundaries of AI to enhance creativity and social connectivity, inviting talented individuals to join us in shaping the future of digital interaction.
About The Role
We are seeking a highly skilled Research / ML Engineer to join our Speech Team. In this role, you will be responsible for developing state-of-the-art speech systems from data specifications through to production deployment. Your primary focus will be on advancing models related to text-to-speech (TTS), voice cloning, controllable TTS, voice conversion, and related tasks. You will play a crucial role in driving the model evaluation cycle, partnering closely with research, data, and infrastructure teams to deliver reliable, efficient, and scalable AI solutions. This position offers a unique opportunity to operate at the intersection of cutting-edge research and practical engineering, contributing to the development of safe, controllable, and trustworthy speech AI systems that can be used at scale.
Qualifications
- Exceptional research and development experience with large-scale audio models (>3 billion parameters, >500,000 hours of data)
- Deep understanding and hands-on experience with transformer architectures, diffusion models, and audio language modeling
- Proficiency in multi-node and multi-GPU distributed training
- Strong software engineering skills with a proven track record of building complex, reliable systems
- Expertise in PyTorch, CUDA, Triton, C++, and performance optimization techniques
- Experience shipping large-scale speech/audio models to production environments
- Background in working with large-scale ML datasets, including data curation, augmentation, and labeling strategies
- Ability to design and execute scientific experiments for model evaluation and improvement
- Notable publications or open-source contributions in speech, audio, or ML domains
- Experience with voice cloning, speech control, and voice generation technologies is preferred
- Strong analytical skills with the ability to iterate on data and triangulate quality signals
Reasons to use Rodeo
I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?
Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.
Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.
Start with a chat, not a search bar
Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.
Graduate Consultant — 2026 Scheme
Why you're a good match
StrongYour economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.
See breakdownIt searches the market for you
Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.
Why you're a good match
You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.
Experience fit
Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.
Only hits
No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.
Responsibilities
- Architect, implement, pre-train, fine-tune, and post-train large-scale speech models, including alignment techniques
- Lead small research projects independently while collaborating on larger team initiatives
- Design, run, and analyze experiments to enhance model performance and understanding
- Develop and improve development tools to increase team productivity and efficiency
- Contribute across the entire model development stack, from low-level optimizations to high-level design
- Define data requirements and collaborate on data acquisition, curation, augmentation, and synthetic data strategies
- Design automated evaluation metrics, including subjective listening tests and objective benchmarks such as SV/WER/ASR metrics
- Harden the training, evaluation, and inference pipelines, ensuring robustness, low latency, and cost-efficiency
- Partner with infrastructure teams to scale training and inference on cloud platforms, ensuring model reliability and observability
- Contribute to safety, consent, and misuse mitigation measures to promote responsible speech technology development


Get help with your application
Your very own career expert that helps elevate your application to the next level.
Benefits
- Competitive salary and generous equity options
- Comprehensive medical, dental, and vision insurance coverage, with most premiums covered by Cantina
- Paid time off including 15 PTO days, 10 sick days, 15 holidays, and 2 floating holidays
- Generous parental leave and fertility support programs
- 401(k) retirement savings plan
- Lifestyle spending account of $500 per month for personal use
- Complimentary lunch and snacks for in-office employees
- Membership to One Medical and additional wellness benefits
Equal Opportunity
Cantina is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, or veteran status. We believe in fostering a workplace where everyone can thrive and contribute to our mission of redefining social interaction through AI technology.
“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”
Jessica, London
Skills
Location