This project ingests movie data from TMDB, stores raw JSON in Google Cloud Storage, transforms it with PySpark, loads curated parquet data into BigQuery, and builds analytics models with dbt.