> Markdown version of [/jobs/ext/2250359-data-lineage-governance-analyst](https://www.wearedevelopers.com/jobs/ext/2250359-data-lineage-governance-analyst). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Lineage & Governance Analyst - **Company:** Boundaryless LTD - **Location:** London, UK - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Airflow, Amazon S3, Data Analysis, Apache HTTP Server, CA Workload Automation Ae, Big Data, Cloudera Impala, Continuous Integration, Data Definition Language, Data Dictionary, Information Engineering, Data Governance, Data Infrastructure, Extract Transform Load (ETL), Data Mapping, Data Transformation, Distributed Computing Environment, Hadoop Distributed File System, Apache Hive, IBM InfoSphere (ETL Tools), Python (Programming Language), Meta-Data Management, Metadata Standards, Nagios, OpenShift, Oracle Databases, Scrum Methodology, SQL Databases, Tableau (Software), Parquet, Sql Optimization, Apache Spark, Git, Microsoft Fabric, Data Lakes, Pyspark, Data Lineage, Collibra, Data Analytics, Data Pipelines - **Published:** August 26, 2026 - **Apply:** https://www.collegerecruiter.com/job/2815168001-data-lineage--governance-analyst ## About the Role * Minimum 3 years of relevant experience in data analytics, data quality, reporting controls, or data transformation programs (preferably in financial services)., * Minimum 3 years of relevant experience in data governance, lineage, metadata management, or controls programs within finance/banking. * Strong understanding of data lineage concepts: technical lineage, business lineage, column-level lineage, impact analysis, and provenance. * Hands-on experience with data lineage / metadata tooling in enterprise environments (e.g., Collibra, Alation, Informatica EDC/IDMC, IBM Infosphere, Microsoft Purview, Apache Atlas, Amundsen, DataHub or similar). * Proven ability to build lineage for complex platforms: data lakes, warehouses, marts, and distributed processing (Spark-based pipelines). * Strong proficiency in SQL for tracing transformations and validating mappings across layers. * Working knowledge of ETL/ELT patterns, data modeling (dimensional + normalized), and batch scheduling dependencies. * Ability to interpret data transformation logic from pipelines (Spark SQL / PySpark / Hive queries / orchestration configs). * Strong documentation capability: source-to-target mappings, lineage diagrams, data dictionaries, metadata standards, and control evidence packs. Technical Skills * Strong proficiency in Python (data analysis/automation for metadata extraction, validation scripts, rule checks). * Hands-on experience with PySpark and Spark SQL in production environments. * Solid knowledge of Hive, Impala, HDFS, and Parquet. * Advanced SQL skills; experience with Oracle databases is preferred. * Working knowledge of Autosys & Apache Airflow. * Experience with CI/CD tools (Git, Harness, UrbanCode Deploy (UCD), Red Hat OpenShift). * Familiarity with AWS S3 for large-scale data storage. * Exposure to Tableau (understanding data sources, extracts, dependencies) is a plus., * Experience with regulatory reporting data domains (risk, liquidity, capital, finance, BCBS 239 alignment, etc.). * Knowledge of data governance operating models: CDEs, data ownership, stewardship, data quality dimensions. * Experience creating audit-ready documentation and participating in audit walkthroughs. * Experience working in Agile/Scrum delivery models. * Familiarity with monitoring and alerting tools for data pipelines. ## Description Role Description * The Technical Analyst - Data Lineage will support a Data Governance, Controls, and Reporting program for a top-tier banking client. * Responsible for establishing and validating end-to-end lineage across critical datasets used in operational and regulatory reporting. * Translate governance and reporting requirements into actionable lineage deliverables (source-to-target mapping, lineage diagrams, metadata standards, and audit evidence). * Work with Data Platform, Data Engineering, Architecture, Risk, Compliance, and Security teams to define lineage standards, metadata capture, and control points. * Maintain lineage artifacts for Critical Data Elements (CDEs), key reports, and priority data products. * Support control design to ensure traceability from source systems * transformations * curated layers * consumption (dashboards/reports/APIs). * Actively participate from discovery workshops through to implementation and continuous improvement. * Ensure traceability from data definition * transformation logic * lineage evidence * audit readiness. Location * The role supports one of our top-tier banking clients in London (Canary Wharf) and requires a minimum of three days on-site presence. * This is a permanent position based in the UK. We will only consider applicants who are eligible to work in the UK. For this role do NOT offer visa sponsorship. Experience Requirements & Qualifications * Minimum 3 years of relevant experience in data analytics, data quality, reporting controls, or data transformation programs (preferably in financial services). Core Skills & Experience * Minimum 3 years of relevant experience in data governance, lineage, metadata management, or controls programs within finance/banking. * Strong understanding of data lineage concepts: technical lineage, business lineage, column-level lineage, impact analysis, and provenance. * Hands-on experience with data lineage / metadata tooling in enterprise environments (e.g., Collibra, Alation, Informatica EDC/IDMC, IBM Infosphere, Microsoft Purview, Apache Atlas, Amundsen, DataHub or similar). * Proven ability to build lineage for complex platforms: data lakes, warehouses, marts, and distributed processing (Spark-based pipelines). * Strong proficiency in SQL for tracing transformations and validating mappings across layers. * Working knowledge of ETL/ELT patterns, data modeling (dimensional + normalized), and batch scheduling dependencies. * Ability to interpret data transformation logic from pipelines (Spark SQL / PySpark / Hive queries / orchestration configs). * Strong documentation capability: source-to-target mappings, lineage diagrams, data dictionaries, metadata standards, and control evidence packs. Technical Skills * Strong proficiency in Python (data analysis/automation for metadata extraction, validation scripts, rule checks). * Hands-on experience with PySpark and Spark SQL in production environments. * Solid knowledge of Hive, Impala, HDFS, and Parquet. * Advanced SQL skills; experience with Oracle databases is preferred. * Working knowledge of Autosys & Apache Airflow. * Experience with CI/CD tools (Git, Harness, UrbanCode Deploy (UCD), Red Hat OpenShift). * Familiarity with AWS S3 for large-scale data storage. * Exposure to Tableau (understanding data sources, extracts, dependencies) is a plus. Nice-to-Have * Experience with regulatory reporting data domains (risk, liquidity, capital, finance, BCBS 239 alignment, etc.). * Knowledge of data governance operating models: CDEs, data ownership, stewardship, data quality dimensions. * Experience creating audit-ready documentation and participating in audit walkthroughs. * Experience working in Agile/Scrum delivery models. * Familiarity with monitoring and alerting tools for data pipelines. Experience Requirements & Qualifications * Conduct discovery workshops to identify priority reports, data products, and Critical Data Elements (CDEs). * Build and maintain end-to-end lineage across systems, including column-level mappings where required. * Produce and maintain Source-to-Target Mapping (STTM) documentation and metadata standards. * Validate lineage accuracy by tracing logic through SQL/Spark transformations and pipeline configurations. * Support impact analysis for proposed changes (upstream/downstream dependencies, report impact, control impact). * Partner with engineers and platform teams to improve metadata capture and lineage automation (where possible). * Define lineage-related control points and produce audit-ready evidence (diagrams, mappings, query proofs, run evidence). * Support UAT by validating that reported numbers can be traced and explained back to trusted sources. * Maintain the lineage backlog and track changes across releases to ensure artifacts remain current. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Best Companies to work for in London: Top 25 Companies in 2023](https://www.wearedevelopers.com/magazine/187-best-companies-to-work-for-in-london-top-25-companies-in-2023)