Rodeo
ResourcesPartnersSign in

Boundaryless Automation

Data Lineage & Governance Analyst

London
Posted about 9 hours ago
Sign up to applySee more jobs like this

How your CV stacks up

1Upload CV
2Analyse CV
3Improve CV

Upload your CV to see how well it fits this job role

?%

Role Description

The Technical Analyst – Data Lineage will support a Data Governance, Controls, and Reporting program for a top-tier banking client.

  • Responsible for establishing and validating end-to-end lineage across critical datasets used in operational and regulatory reporting.
  • Translate governance and reporting requirements into actionable lineage deliverables (source-to-target mapping, lineage diagrams, metadata standards, and audit evidence).
  • Work with Data Platform, Data Engineering, Architecture, Risk, Compliance, and Security teams to define lineage standards, metadata capture, and control points.
  • Maintain lineage artifacts for Critical Data Elements (CDEs), key reports, and priority data products.
  • Support control design to ensure traceability from source systems → transformations → curated layers → consumption (dashboards/reports/APIs).
  • Actively participate from discovery workshops through to implementation and continuous improvement.
  • Ensure traceability from data definition → transformation logic → lineage evidence → audit readiness.

Location

The role supports one of our top-tier banking clients in London (Canary Wharf) and requires a minimum of three days on-site presence. This is a permanent position based in the UK. We will only consider applicants who are eligible to work in the UK. For this role, we do NOT offer visa sponsorship.

Experience Requirements & Qualifications

  • Minimum 3 years of relevant experience in data analytics, data quality, reporting controls, or data transformation programs (preferably in financial services).

Reasons to use Rodeo

I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?

Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.

Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.

Start with a chat, not a search bar

Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.

P

Graduate Consultant — 2026 Scheme

PwC·London, UK
£35,000/yr

Why you're a good match

Strong

Your economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.

See breakdown
Save jobNot relevant
View details

It searches the market for you

Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.

Why you're a good match

You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.

See breakdown
Strong

Experience fit

Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.

See breakdown
Strong

Only hits

No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.

Core Skills & Experience

  • Minimum 3 years of relevant experience in data governance, lineage, metadata management, or controls programs within finance/banking.
  • Strong understanding of data lineage concepts: technical lineage, business lineage, column-level lineage, impact analysis, and provenance.
  • Hands-on experience with data lineage / metadata tooling in enterprise environments (e.g., Collibra, Alation, Informatica EDC/IDMC, IBM Infosphere, Microsoft Purview, Apache Atlas, Amundsen, DataHub or similar).
  • Proven ability to build lineage for complex platforms: data lakes, warehouses, marts, and distributed processing (Spark-based pipelines).
  • Strong proficiency in SQL for tracing transformations and validating mappings across layers.
  • Working knowledge of ETL/ELT patterns, data modeling (dimensional + normalized), and batch scheduling dependencies.
  • Ability to interpret data transformation logic from pipelines (Spark SQL / PySpark / Hive queries / orchestration configs).
  • Strong documentation capability: source-to-target mappings, lineage diagrams, data dictionaries, metadata standards, and control evidence packs.

Technical Skills

  • Strong proficiency in Python (data analysis/automation for metadata extraction, validation scripts, rule checks).
  • Hands-on experience with PySpark and Spark SQL in production environments.
  • Solid knowledge of Hive, Impala, HDFS, and Parquet.
  • Advanced SQL skills; experience with Oracle databases is preferred.
  • Working knowledge of Autosys & Apache Airflow.
  • Experience with CI/CD tools (Git, Harness, UrbanCode Deploy (UCD), Red Hat OpenShift).
  • Familiarity with AWS S3 for large-scale data storage.
  • Exposure to Tableau (understanding data sources, extracts, dependencies) is a plus.

Get help with your application

Your very own career expert that helps elevate your application to the next level.

Get help applying for this job

Nice-to-Have

  • Experience with regulatory reporting data domains (risk, liquidity, capital, finance, BCBS 239 alignment, etc.).
  • Knowledge of data governance operating models: CDEs, data ownership, stewardship, data quality dimensions.
  • Experience creating audit-ready documentation and participating in audit walkthroughs.
  • Experience working in Agile/Scrum delivery models.
  • Familiarity with monitoring and alerting tools for data pipelines.

Main Tasks and Responsibilities

  • Conduct discovery workshops to identify priority reports, data products, and Critical Data Elements (CDEs).
  • Build and maintain end-to-end lineage across systems, including column-level mappings where required.
  • Produce and maintain Source-to-Target Mapping (STTM) documentation and metadata standards.
  • Validate lineage accuracy by tracing logic through SQL/Spark transformations and pipeline configurations.
  • Support impact analysis for proposed changes (upstream/downstream dependencies, report impact, control impact).
  • Partner with engineers and platform teams to improve metadata capture and lineage automation (where possible).
  • Define lineage-related control points and produce audit-ready evidence (diagrams, mappings, query proofs, run evidence).
  • Support UAT by validating that reported numbers can be traced and explained back to trusted sources.
  • Maintain the lineage backlog and track changes across releases to ensure artifacts remain current.
Trusted by 25,000+ job seekers

“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”

Jessica, London

Get help applying for this job

Skills

Data Lineage
Data Governance
Metadata Management
SQL
Python
PySpark
Spark SQL
Collibra
Alation
Informatica EDC/IDMC
AWS S3
Apache Airflow
Source-to-Target Mapping
Data Modeling
ETL/ELT
Impact Analysis

Location

London, England, United Kingdom

Sign up to applySee more jobs like this