Lineage
Overview
Lineage information is critical for metadata systems. Gravitino supports data lineage by leveraging OpenLineage and provides a specific Spark JAR to collect lineage information with the Gravitino identifier. For details, see the Gravitino Spark lineage page. Additionally, the Gravitino server provides a lineage process framework to receive, process, and sink OpenLineage events to other systems.
Capabilities
- Supports column lineages.
- Supports lineage across diverse Gravitino catalogs like fileset, Iceberg, Hudi, Paimon, Hive, Model, etc.
- Supports Spark.