Data Governance & Metadata
Unified Metadata Management
★ 4.0
Open Data Management System
★ 4.1
pip install apache-gravitinopip install ckanapipip install apache-gravitinopip install ckanapiPython data engineers use Gravitino's REST API to register and discover table schemas centrally when working across multiple compute engines — registering an Iceberg table in Gravitino makes it discoverable to Spark, Trino, and Flink without duplicating schema definitions. Python scripts automate schema registration after new pipeline outputs are created.
Python data engineers use the `ckanapi` library to programmatically harvest open datasets from CKAN portals — listing available datasets, downloading CSV or JSON resources, and ingesting them into internal pipelines. Government open data platforms (data.gov, data.gov.uk) run on CKAN, making it the standard entry point for public data ingestion workflows.
Data Governance & Metadata
Amundsen vs Apache Atlas
Data Governance & Metadata
Apache Atlas vs CKAN
Data Governance & Metadata
Apache Atlas vs Marquez
Data Governance & Metadata
Apache Atlas vs DataHub
Data Governance & Metadata
Apache Atlas vs Collibra
Data Governance & Metadata
Apache Atlas vs Apache Gravitino
Individual Tool Pages