Rodeo
Get started

Women in Data®

AI Data Engineer

London
Posted about 18 hours ago
Sign up to applySee more jobs like this
Get notified of more jobs like this · No spam, ever

How your CV stacks up

1Upload CV
2Analyse CV
3Improve CV

Upload your CV to see how well it fits this job role

?%

Locations: London, Bristol, Newcastle upon Tyne, Birmingham, Manchester

Who You Will Be Working With

As a Data Engineer at Capgemini, you'll design, build and operate reliable, scalable data pipelines and data products: the trusted foundations that power analytics, AI, GenAI and increasingly agentic systems. You'll work hands-on with modern data engineering tooling and cloud services to ingest, transform and serve data, and you'll prepare and govern the datasets that AI models and AI agents depend on to behave safely and correctly. You'll also use AI-assisted engineering to accelerate your own delivery, while keeping clear human ownership of every outcome.

You'll be part of the Data Platforms team within the Insights and Data Global Practice, which has seen strong, sustained growth across a wide range of sectors. Data Platforms is home to Data Engineers, Platform Engineers, Solution Architects and Business Analysts driving our customers' digital and data transformation on modern cloud platforms. We specialise in the latest frameworks, reference architectures and technologies across AWS, Azure and GCP, and data platforms such as Databricks and Snowflake.

PLEASE NOTE: Security Clearance: To be successfully appointed to this role, you must be eligible to obtain Security Check (SC) clearance or DV (Developed Vetting clearance). To obtain SC clearance, the successful applicant must have resided continuously within the United Kingdom for the last 5 years, along with other criteria and requirements.

Throughout the recruitment process, you will be asked questions about your security clearance eligibility such as, but not limited to, country of residence and nationality.

Some posts are restricted to sole UK Nationals for security reasons; therefore, you may be asked about your citizenship in the application process.

The Focus of Your Role

You'll build and run the data products that modern AI depends on. That means classic, high-quality data engineering (batch and streaming pipelines, lakehouse and warehouse patterns, well-governed and auditable datasets), and the newer discipline of engineering data for AI and agents: retrieval pipelines, embeddings and vector stores, feature and context preparation, and the observability needed when the systems consuming your data are non-deterministic and can't simply be unit tested.

What You’ll Be Doing:

  • You'll work within agreed security and compliance boundaries throughout, which matters in the regulated and public sector environments we operate in.
  • Build and maintain data pipelines and data products. Use appropriate ETL/ELT and distributed processing to ingest, transform and curate trusted data in cloud storage and analytics platforms.
  • Engineer data for AI and agentic use cases. Prepare curated, governed datasets and features for analytics and ML; build retrieval and embedding pipelines (vector stores, chunking, metadata) that serve RAG and agent workflows; and help expose data to AI systems through patterns such as APIs and the Model Context Protocol (MCP).
  • Apply data modelling, quality and governance. Develop scalable models, apply validation and quality checks, and maintain lineage and documentation so data products are reliable and auditable. This matters even more when an AI agent, not just a human, is acting on them.
  • Build in observability for non-deterministic systems. Implement logging, alerting and SLAs for production pipelines, and contribute to evaluation and monitoring of AI-facing data and outputs (e.g. drift, retrieval quality, anomaly detection) with clear human review.
  • Use AI-assisted engineering responsibly. Accelerate development, testing and documentation with coding assistants, applying good prompt hygiene, protecting confidential data, and owning the results.
  • Collaborate across teams. Work with business analysts, platform engineers, data scientists and DevOps to deliver secure, well-tested solutions in an agile environment, communicating your design decisions clearly.
  • Keep raising the bar. Use version control, code review, automated testing and CI/CD; share knowledge, contribute to accelerators and standards, and keep current with modern data and AI engineering practice.

Reasons to use Rodeo

I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?

Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.

Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.

Start with a chat, not a search bar

Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.

P

Graduate Consultant — 2026 Scheme

PwC·London, UK
£35,000/yr

Why you're a good match

Strong

Your economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.

See breakdown
Save jobNot relevant
View details

It searches the market for you

Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.

Why you're a good match

You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.

See breakdown
Strong

Experience fit

Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.

See breakdown
Strong

Only hits

No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.

What You Will Bring and Experience Needed

You'll bring solid, hands-on experience delivering data engineering in production, and a genuine interest in how AI and agentic systems change what good data engineering looks like. You don't need to have done all of the AI-specific work below already, but you should be eager to, and able to show the engineering fundamentals that make it possible.

Essential:

  • Hands-on experience delivering data pipelines and data platforms in production environments.
  • Strong Python and SQL, with sound software engineering practice (Git, code review, unit testing, CI/CD) and the ability to troubleshoot production issues.
  • Experience with distributed data processing and modern lakehouse/warehouse patterns, with good data modelling and performance instincts.
  • Experience with cloud data services (storage, compute and orchestration) on at least one major cloud.
  • Practical use of AI coding assistants in a real engineering workflow, with an understanding of data confidentiality and secure, responsible use.
  • A continuous-learning mindset and clear communication with technical and non-technical colleagues.

Nice to have:

  • Exposure to GenAI/agentic building blocks: RAG, embeddings and vector search, LLM orchestration (e.g. LangGraph, LlamaIndex, Semantic Kernel), or LLM evaluation/observability (e.g. LangSmith, Ragas).
  • Streaming and event-driven experience (e.g. Kafka, Spark Structured Streaming).
  • Relevant cloud and/or data engineering certifications.

Specialist tracks (we'd love depth in one of these; you don't need all of them):

  • Azure / Databricks: Azure Data Lake Storage, Databricks, Apache Spark, Delta Lake, Azure Data Factory, MLflow; and, for the AI edge, Databricks Mosaic AI, Vector Search and Unity Catalog, or Azure OpenAI, Azure AI Foundry and Azure AI Search.
  • AWS: Glue, Lambda, Step Functions, Kinesis, EMR, Athena, Redshift, S3 data lakes; and, for the AI edge, Amazon Bedrock, Knowledge Bases for Bedrock

Get help with your application

Your very own career expert that helps elevate your application to the next level.

Get help applying for this job

Additional Info

  • Hybrid working: The places that you work from day to day will vary according to your role, your needs, and those of the business; it will be a blend of Company offices, client sites, and your home; noting that you will be unable to work at home 100% of the time.
  • If you are successfully offered this position, you will go through a series of pre-employment checks, including identity, nationality (single or dual) or immigration status, employment history going back 3 continuous years, and unspent criminal record check (known as Disclosure and Barring Service)

What We’ll Offer You

  • You will be encouraged to have a positive work-life balance. Our hybrid-first way of working means we embed hybrid working in all that we do and make flexible working arrangements the day-to-day reality for our people. All UK employees are eligible to request flexible working arrangements.
  • You will be empowered to explore, innovate, and progress. You will benefit from Capgemini’s ‘learning for life’ mindset, meaning you will have countless training and development opportunities from thinktanks to hackathons, and access to 250,000 courses with numerous external certifications from AWS, Microsoft, Harvard Manage Mentor, Cybersecurity qualifications and much more.

Why We’re Different

At Capgemini, we help organisations across the world become more agile, more competitive, and more successful. Smart, tailored, often ground-breaking technical solutions to complex problems are the norm. But so, too, is a culture that’s as collaborative as it is forward thinking. Working closely with each other, and with our clients, we get under the skin of businesses and to the heart of their goals. You will too.

Capgemini is proud to represent nearly 130 nationalities and its cultural diversity. Our holistic definition of diversity extends beyond gender, gender identity, sexual orientation, disability, ethnicity, race, age, and religion. Capgemini views diversity as everything that makes us who we are as an organization, including our social background, our experiences in life and work, our communication styles and even our personality. These dimensions contribute to the type of diversity we value the most: diversity of thought.

Women in Data® advertise roles on behalf of our partners, alliances, and members.

We are proud supporters of Women in Data. Connect, engage and belong to the largest free female data community in the UK – visit: www.womenindata.co.uk to join our community.

Stay connected! Follow us on LinkedIn for updates on career opportunities and more.

Trusted by 25,000+ job seekers

“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”

Jessica, London

Get help applying for this job

Location

London, England, United Kingdom

Sign up to applySee more jobs like this