End-to-end data pipeline that ingests YouTube Live Stream data using the YouTube API and stores it as Parquet in S3. Utilizes AWS Lambda, SNS, and Redshift for automated data loading and processing, orchestrated via Databricks.
-
Updated
Jun 3, 2025 - Python