What problem does it solve?
This Skill removes the need for manual data replication or custom integration work when you need to query data stored in remote cross-cloud lakehouse catalogs like Databricks Unity Catalog or AWS Glue Data Catalog directly from GCP analytics tools like BigQuery and Spark.
Core Features & Use Cases
- Cross-Catalog Federation: Supports connecting GCP BigLake Iceberg federated catalogs to remote Databricks Unity Catalog and AWS Glue Data Catalog instances hosted on either GCP or AWS.
- Native GCP Querying: Lets you run queries on remote catalog data directly from GCP services without moving or copying the underlying data to GCP.
- Real-World Use Case: If your e-commerce team stores customer order data in a Databricks Unity Catalog on AWS and wants to run monthly sales performance reports in BigQuery without replicating the entire dataset to GCP, this skill automates the full end-to-end federation setup.
Quick Start
Use the federate-lakehouse-catalog skill to set up a secure federated connection between your GCP project and your AWS-hosted AWS Glue Data Catalog to query your inventory data from BigQuery.