Skip to main content Link Menu Expand (external link) Document Search Copy Copied

Task 02: Explore the data in the Azure Databricks environment with Unity Catalog

Introduction

In Task 1, you saw how Zava uses DLT pipelines to create a Medallion architecture for their data. Now let’s look at how Zava manages data governance by using Unity Catalog.

Key steps

  1. In the left pane, select Catalog.

  2. In the list of catalogs, expand dbkwks@lab.LabInstance.Id. Then, expand schema@lab.LabInstance.Id_processeddata.

  3. Verify that there are tables created by the pipeline are listed. Refresh the page if you don’t see the tables.

ghnopstd.jpg

  1. Select gold_country_wise_revenue.

  2. On the command bar for the table, select Lineage.

    5278ahyz.jpg

  3. Select See lineage graph to view the full upstream and downstream data flow.

    9hre8hsx.jpg

    View the visualization that shows how the Gold table is derived from upstream Bronze and Silver tables.

    f9hpndof.jpg

  4. Close the lineage graph page. Then, on the command bar for the table, select Overview.

    dtkay598.jpg

  5. On the Overview tab, select AI generate to automatically create column descriptions by using Azure Databricks data intelligence.

    c2k40jev.jpg

  6. Review the descriptions that the system generated and then select Save all.

    1qnqeh23.jpg