I created a Chart based on the Hive data (default....
# ingestion
b
I created a Chart based on the Hive data (default.table) in Superset, and when I ingested from Hive and Superset to DataHub, and Lineage was generated. (great!šŸ˜†) However, since the urn of the original dataset is different between the dataset directly ingested from Hive and the Chart ingested from Superset, the dataset was registered as a duplicate.šŸ¤” • dataset generated by Hive ingestion: ā—¦ urnlidataset(urnlidatasetPlatform:hive,default.table,PROD) • dataset generated by Superset ingestion: ā—¦ urnlidataset(urnlidatasetPlatform:hive,_*Apache Hive.*_default.table,PROD) The urn of the dataset ingested from Superset will not match because the database name set in Superset is added to the urn. Is it possible to match urns, for example adding "Apache Hive" to the urn ingested from Hive and ingest it?
Sorry. I solved it. I was able to match the urn by setting database_alias in superset's recipe. In my case, I wrote the following.
Copy code
database_alias: 
  Apache Hive: ""