Real-time Data Warehouse with Apache Flink & Apache Kafka & Apache Hudi
☆121Dec 15, 2023Updated 2 years ago
Alternatives and similar repositories for Real-time-Data-Warehouse
Users that are interested in Real-time-Data-Warehouse are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This project demonstrates Real-Time streaming of CDC data from MySql to Apache Iceberg using Flink SQL Client for faster data analytics a…☆25Jan 16, 2024Updated 2 years ago
- Traditionally, engineers were needed to implement business logic via data pipelines before business users can start using it. Using this …☆12Updated this week
- ☆18Nov 26, 2024Updated last year
- ☆11Jul 28, 2026Updated last month
- 汇总Apache Hudi中的一些Demo,便于快速上手Apache Hudi(Apache Hudi Demos to help beginners know about Hudi)☆74Sep 13, 2020Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A custom end-to-end analytics platform for customer churn☆10May 15, 2025Updated last year
- The Apache Flink SQL Cookbook is a curated collection of examples, patterns, and use cases of Apache Flink SQL. Many of the recipes are c…☆915Jan 12, 2026Updated 7 months ago
- A repository used in a NiFi Registry demo☆13Mar 11, 2020Updated 6 years ago
- adidas Data Mesh implementation☆12May 13, 2022Updated 4 years ago
- Examples for using Apache Flink® with DataStream API, Table API, Flink SQL and connectors such as MySQL, JDBC, CDC, Kafka.☆65Sep 26, 2023Updated 2 years ago
- ☆176Sep 5, 2023Updated 2 years ago
- Simple akka cluster example.☆12Mar 13, 2015Updated 11 years ago
- 基于Mysql 协议的数据库中间件,支持SQL查询Mysql、Oracle、Clickhouse、Excel、elasticsearch,基于calcite进行SQL解析,并对SQL进行扩展☆14Aug 5, 2024Updated 2 years ago
- ☆16May 1, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for Apache Hudi, Apache Iceberg and Delta Lake analysis☆10Feb 2, 2024Updated 2 years ago
- Apache flink☆19May 15, 2026Updated 3 months ago
- Low Cost, Simple and Scalable Way of Data Replication to Apache Iceberg/Cloud/Data Lake☆327Updated this week
- Apache Airflow advanced functionalities examples☆21Mar 22, 2024Updated 2 years ago
- This plugin provides a useful feature for multi-language☆14Jul 15, 2022Updated 4 years ago
- This project shows how to capture changes from postgres database and stream them into kafka☆41May 17, 2024Updated 2 years ago
- 基于flink的实时流计算web平台☆1,858Dec 2, 2025Updated 8 months ago
- AI 时代的智能数据库☆221Nov 9, 2023Updated 2 years ago
- The Internals of Apache Kafka☆59Dec 19, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆10May 5, 2022Updated 4 years ago
- Adapter for dbt that executes dbt pipelines on Apache Flink☆102Mar 19, 2024Updated 2 years ago
- My Dota 2 Bot Script☆11Jun 6, 2022Updated 4 years ago
- 汇总Apache Hudi相关资料☆556Mar 31, 2026Updated 5 months ago
- This repository hosts materials for the Docker for Data Engineers workshop, offering hands-on exercises and resources tailored for data e…☆16May 23, 2024Updated 2 years ago
- Gitbook Repo for Practical Data Pipeline☆25Feb 4, 2022Updated 4 years ago
- Local AWS EMR - A local service that imitates AWS EMR☆27Jul 5, 2023Updated 3 years ago
- A Picture Management software using MFC☆10Sep 16, 2013Updated 12 years ago
- Example applications in Java, Python and SQL for Kinesis Data Analytics, demonstrating sources, sinks, and operators.☆147May 21, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A data generator source connector for Flink SQL based on data-faker.☆238Jul 24, 2023Updated 3 years ago
- Apache Amoro(incubating) is a Lakehouse management system built on open data lake formats.☆1,171Updated this week
- SDK in python for datarangers products☆17Apr 30, 2026Updated 4 months ago
- ☆116Apr 21, 2023Updated 3 years ago
- Instant access to the Spark cluster from anywhere☆15Nov 10, 2020Updated 5 years ago
- This project provides a reverse proxy for Spark UI on Kubernetes☆16Oct 12, 2023Updated 2 years ago
- Upserts, Deletes And Incremental Processing on Big Data.☆6,229Updated this week