A complete data engineering project demonstrating modern data stack practices with Apache Flink, Iceberg, Trino and Superset
☆27Sep 29, 2025Updated 11 months ago
Alternatives and similar repositories for apache_flink_and_iceberg
Users that are interested in apache_flink_and_iceberg are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This project demonstrates Real-Time streaming of CDC data from MySql to Apache Iceberg using Flink SQL Client for faster data analytics a…☆25Jan 16, 2024Updated 2 years ago
- 📡 Real-time data pipeline with Kafka, Flink, Iceberg, Trino, MinIO, and Superset. Ideal for learning data systems.☆79Jan 18, 2025Updated last year
- Arrow-Powered Data Exchange☆15Feb 7, 2025Updated last year
- ☆14Jun 10, 2024Updated 2 years ago
- Apache Hive Metastore in Standalone Mode With Docker☆14Jul 22, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Trino ODBC driver☆15Jan 10, 2024Updated 2 years ago
- An Ibis back-end for the GizmoSQL Arrow Flight SQL Server (with the DuckDB engine)☆18Aug 24, 2026Updated last week
- ☆13Jan 31, 2024Updated 2 years ago
- A data transfer tool using ADBC and Go built for your Agent to use☆23Jul 1, 2026Updated 2 months ago
- A Cheap Alternative to AWS Athena: Lambda x DuckDB☆16Mar 22, 2023Updated 3 years ago
- Apache iceberg Spark s3 examples☆20Mar 1, 2024Updated 2 years ago
- 🌟 Examples of use cases that utilize Decodable, as well as demos for related open-source projects such as Apache Flink, Debezium, and Po…☆93Jun 20, 2025Updated last year
- version 2 of the Unified Cybersecurity Ontology☆16May 7, 2017Updated 9 years ago
- Trino AI SQL Functions☆15Nov 29, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A cli for spinning up and managing Ray clusters for the Daft Query Engine.☆14Feb 15, 2025Updated last year
- Spark* plug-in for accelerating Spark* SQL performance by using cache and index at SQL data source layer.☆37Jan 3, 2023Updated 3 years ago
- Alerting and monitoring tool for Apache Spark☆23May 20, 2022Updated 4 years ago
- End-to-end data platform leveraging the Modern data stack☆52Apr 10, 2024Updated 2 years ago
- BigQuery Schema Conversion Tool☆24Oct 6, 2020Updated 5 years ago
- Expert knowledge skills for Claude Code to help with specific technologies and tools.☆37Jul 29, 2026Updated last month
- Применение Debezium для обработки потоковых данных: Основные концепции, примеры.☆24Apr 12, 2025Updated last year
- Material for a course on applied machine-learning for scientists. Taught at EPFL in spring 2018.☆11May 3, 2018Updated 8 years ago
- A set of transformations for Kafka Connect☆23Mar 1, 2026Updated 6 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Go library for decoding generic map values and native Go structures into Arrow.☆19Jan 30, 2026Updated 7 months ago
- dbc is the command-line tool for installing and managing ADBC drivers☆131Updated this week
- Olympia is a storage-only open catalog format for big data analytics, ML & AI.☆16May 5, 2025Updated last year
- Extension for Greenplum to read Apache Iceberg☆18Updated this week
- ☆25Feb 7, 2024Updated 2 years ago
- claude-code generated parquet metadata vizualizer that runs in your browser☆16Dec 8, 2025Updated 8 months ago
- Unity Catalog Explorer is a TypeScript + Next.js based Web UI for the Unity Catalog OSS.☆13Jun 29, 2024Updated 2 years ago
- Go ADBC driver for DuckDB's Quack remote protocol (quack:// URI scheme). Returns Apache Arrow RecordBatches; supports bulk-ingest via APP…☆28Aug 26, 2026Updated last week
- A modern theme for JSON Resume which is self-contained. The content of the resume will work offline and can be hosted without depending o…☆34Jun 27, 2026Updated 2 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Apache Hive Metastore as a Standalone server in Docker☆81Aug 22, 2024Updated 2 years ago
- ☆17Mar 19, 2024Updated 2 years ago
- Go library to stream Kafka protobuf messages to DuckDB☆26Mar 31, 2026Updated 5 months ago
- dbt-databend adapter plugin☆10May 30, 2024Updated 2 years ago
- 数据仓库实战:Hive、HBase、Kylin、ClickHouse☆23May 13, 2026Updated 3 months ago
- Jido implementation of Managed Agents on Phoenix☆23Updated this week
- Java implementation for performing operations on Apache Iceberg and Hive tables☆22May 25, 2026Updated 3 months ago