Data lineage tracking is the automated capture and maintenance of provenance metadata as data is created, transformed and moved across pipelines and platforms. It instruments transformations and queries to build a continuously updated lineage graph rather than a static map. Often expressed using provenance vocabularies such as PROV-O, it underpins reproducibility, trust and governance in data fabric architectures.

Content

  • Tracking instruments ETL jobs, SQL engines and orchestration tools to emit lineage events that assemble into a live graph of derivations. This continuous record enables reproducibility of analyses, automated impact assessment and standards-based interchange of provenance.