☆145Jan 31, 2023Updated 3 years ago
Alternatives and similar repositories for Spotify_etl
Users that are interested in Spotify_etl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Apartments Data Pipeline using Airflow and Spark.☆24Mar 28, 2022Updated 4 years ago
- ☆177May 20, 2022Updated 4 years ago
- I am using confluent Kafka cluster to produce and consume scraped data. In this project, I've created a real-time data pipeline that uti…☆29May 2, 2023Updated 3 years ago
- Integrating with Spotify API and extracting Data. Deploying code on AWS Lambda for Data Extraction. Adding trigger to run the extraction …☆12Jul 5, 2023Updated 3 years ago
- Pipeline that extracts data from Crinacle's Headphone and InEarMonitor databases and finalizes data for a Metabase Dashboard. The dashboa…☆270Jan 1, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- In this project, we will build and ETL(Extract,Transform,Load) pipeline using the Spotify API on AWS. The pipeline will retrieve data fro…☆25May 6, 2023Updated 3 years ago
- ☆217Aug 13, 2023Updated 3 years ago
- Orchestrate Modal and OpenAI workloads with Dagster☆13Dec 11, 2024Updated last year
- capstone project for Dataengineer.io bootcamp Public Repo☆12Feb 20, 2024Updated 2 years ago
- Code for "Advanced data transformations in SQL" free live workshop☆94May 5, 2025Updated last year
- Repository for Data Engineering Zoomcamp 2024☆14Mar 25, 2024Updated 2 years ago
- A data engineering project with Kafka, Spark Streaming, dbt, Docker, Airflow, Terraform, GCP and much more!☆918Apr 16, 2022Updated 4 years ago
- ☆19May 27, 2023Updated 3 years ago
- ☆20Apr 3, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Data engineering project using UK Bus Open Data Service (BODS) to calculate late buses in real-time for any selected region in England. P…☆32Apr 2, 2023Updated 3 years ago
- dbt(databuildtool)-masterclass-netflix-project☆18Jul 8, 2025Updated last year
- ☆406Jan 26, 2025Updated last year
- Data Engineering Project in GCP☆23Mar 29, 2023Updated 3 years ago
- Pipeline that extracts data from the Spotify API to build a more detailed version of Spotify Wrapped☆50Sep 8, 2026Updated 2 weeks ago
- Building a Modern Data Lake with Minio, Spark, Airflow via Docker.☆23May 11, 2024Updated 2 years ago
- ☆45Jul 6, 2024Updated 2 years ago
- This project involves an ETL (Extract, Transform, Load) process to analyze sleep data exported from Apple Health☆29Apr 29, 2023Updated 3 years ago
- This project introduces PySpark, a powerful open-source framework for distributed data processing. We explore its architecture, component…☆47Sep 26, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- DataTalks Workshop Materials☆19Mar 18, 2024Updated 2 years ago
- YouTube tutorial project☆112Oct 17, 2023Updated 2 years ago
- Data Engineering with Scala, published by Packt☆28Apr 22, 2026Updated 5 months ago
- Data warehouse tech stack with PostgreSQL, DBT and Airflow☆19Dec 29, 2025Updated 8 months ago
- Master's thesis on Big Data☆36Aug 14, 2022Updated 4 years ago
- Demonstrating the capabilities of DuckDB as a transformation engine for data lakes☆33Oct 8, 2024Updated last year
- ☆13May 13, 2022Updated 4 years ago
- Data pipeline that scrapes Rust cheater Steam profiles☆53Feb 13, 2022Updated 4 years ago
- End-to-end data pipeline that ingests, processes, and stores data. It uses Apache Airflow to schedule scripts that fetch data from an API…☆23Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Deploy a complete data stack in just a couple of minutes.☆15Mar 6, 2024Updated 2 years ago
- Collection of cookiecutter starter templates for streamlit projects☆15Apr 20, 2022Updated 4 years ago
- A batch processing data pipeline, using AWS resources (S3, EMR, Redshift, EC2, IAM), provisioned via Terraform, and orchestrated from loc…☆25May 14, 2022Updated 4 years ago
- Join 15k builders to the Real-World ML Newsletter ⬇️⬇️⬇️☆18Apr 19, 2024Updated 2 years ago
- Kafka Connect: How to create a real time data pipeline using Change Data Capture (CDC)☆13Jan 24, 2021Updated 5 years ago
- ☆12Jul 22, 2025Updated last year
- Learn the entire ETL process based on Spotify API data☆269Feb 1, 2021Updated 5 years ago