Solution
Amagi rebuilt its data platform as the Amagi Data Platform (ADP) on Databricks, consolidating Dataproc, Snowflake, and multiple in-house systems into a single Lakehouse spanning AWS and GCP. Lakeflow Spark Declarative Pipelines ingest content metadata, asset catalogs, advertising data, service streams, program schedules, and user data into governed Delta Lake tables, where Unity Catalog enforces lineage, access controls, sensitive-data tagging, and cross-cloud compliance. Finance, Operations, and Product teams query the same source of truth through Databricks SQL warehouses, Genie natural-language sessions, and notebooks running on serverless compute, while customer-facing data products for streaming platforms, ad-tech partners, and content owners are derived from the same governed Lakehouse, eliminating reconciliation between systems.
Data flow
Source systems — content metadata, asset catalogs, advertising data, service streams, program schedules, and user data — land in Delta Lake via Lakeflow declarative pipelines. Unity Catalog governs lineage, access, and sensitive-data tagging across the multicloud AWS+GCP footprint. Internal users query the same governed Lakehouse through Databricks SQL warehouses, Genie natural-language sessions, and notebooks on serverless compute, while customer-facing data products (streaming analytics, ad attribution, usage and revenue reporting) are derived from the same source of truth.