This project provides an end-to-end data processing and visualization of visa numbers in Japan using PySpark and Plotly. The spark clusters are set up within a Docker container on Azure.
☆11Oct 11, 2023Updated 2 years ago
Alternatives and similar repositories for Japan-visa-data-engineering
Users that are interested in Japan-visa-data-engineering are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An end-to-end data engineering pipeline that fetches real-time YouTube analytics and streams them through Kafka for processing with ksqlD…☆16Sep 19, 2023Updated 2 years ago
- An end-to-end data engineering pipeline that fetches data from Wikipedia, cleans and transforms it with Apache Airflow and saves it on Az…☆31Oct 2, 2023Updated 2 years ago
- This project showcases how to integrate the world of DevOps, focusing on Continuous Integration (CI) and Continuous Deployment (CD) with …☆14Dec 27, 2023Updated 2 years ago
- This repository contains an end-to-end data engineering project using Apache Flink, focused on performing sales analytics. The project de…☆12Nov 18, 2023Updated 2 years ago
- This repository contains the necessary configuration files and DAGs (Directed Acyclic Graphs) for setting up a robust data engineering en…☆25Jan 26, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This project serves as a comprehensive guide to building an end-to-end data engineering pipeline using TCP/IP Socket, Apache Spark, OpenA…☆45Jan 4, 2024Updated 2 years ago
- This project shows how to capture changes from postgres database and stream them into kafka☆41May 17, 2024Updated 2 years ago
- This repository contains the code for a realtime election voting system. The system is built using Python, Kafka, Spark Streaming, Postgr…☆48Dec 11, 2023Updated 2 years ago
- In this project, we setup and end to end data engineering using Apache Spark, Azure Databricks, Data Build Tool (DBT) using Azure as our …☆39Dec 18, 2023Updated 2 years ago
- A data pipeline for processing football data using Python and SQL☆13Sep 12, 2023Updated 2 years ago
- This demo shows how to stream data to cloud databases with Confluent. It includes fully-managed connectors (Oracle CDC, RabbitMQ, MongoDB…☆11Jan 10, 2025Updated last year
- An end-to-end data engineering pipeline that orchestrates data ingestion, processing, and storage using Apache Airflow, Python, Apache Ka…☆339Feb 14, 2025Updated last year
- This repository contains an Apache Flink application for real-time sales analytics built using Docker Compose to orchestrate the necessar…☆51Dec 4, 2023Updated 2 years ago
- ☆16Dec 30, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This project provides a comprehensive data pipeline solution to extract, transform, and load (ETL) Reddit data into a Redshift data wareh…☆227Oct 23, 2023Updated 2 years ago
- This project demonstrates how to use Apache Airflow to submit jobs to Apache spark cluster in different programming laguages using Python…☆48Mar 14, 2024Updated 2 years ago
- ☆16Feb 20, 2026Updated 5 months ago
- Query Iceberg in Trino, Nessie as Catalog, and use minio to replace AWS S3☆27Aug 7, 2025Updated last year
- This project leverages Hadoop, Spark, SQL, and Hive for efficient data integration, transformation, warehousing, and analytics. It provid…☆24Sep 30, 2023Updated 2 years ago
- Udacity Data Engineer Nano Degree - Project-3 (Data Warehouse)☆22Jun 20, 2019Updated 7 years ago
- Repository to host micro service implementation patterns.☆14Jun 25, 2025Updated last year
- GPT-4o Powered Calorie Detecor☆18May 29, 2024Updated 2 years ago
- ☆10Jan 8, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Automatically backing up your Postgres database using NodeJS☆13Nov 14, 2020Updated 5 years ago
- This repository showcases a collection of machine learning projects in various domains, demonstrating my skills and expertise as a data s…☆13Nov 20, 2023Updated 2 years ago
- To gain access, please finish setting up this repository now at:☆35Mar 23, 2026Updated 4 months ago
- ☆30Aug 14, 2025Updated last year
- ☆10Jan 18, 2024Updated 2 years ago
- Realtime Data Engineering Project☆31Jan 12, 2025Updated last year
- ☆43Oct 24, 2024Updated last year
- Transparent sandbox for integration testing against AWS services. Test your infrastructure without changes to your Terraform files or you…☆12Oct 26, 2023Updated 2 years ago
- ☆17Apr 18, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆13May 18, 2026Updated 2 months ago
- ☆17Mar 10, 2023Updated 3 years ago
- ☆12Jan 31, 2026Updated 6 months ago
- Angular JWT refresh token with Interceptor, handle token expiration in Angular 14 - Refresh token before expiration example☆13Sep 20, 2022Updated 3 years ago
- KazeWP is a simple and flexible tool for managing multiple WordPress sites behind a Caddy reverse proxy server. Built with Docker and Bas…☆17Apr 28, 2025Updated last year
- ☆14Mar 11, 2023Updated 3 years ago
- This is an example of using MongoDB as both a source and sink.☆10May 21, 2020Updated 6 years ago