This is a collecton of Amazon CDK projects to show how to directly ingest streaming data from Amazon Mananged Service for Apache Kafka (MSK) and MSK Serverless into Apache Iceberg table in S3 with AWS Glue Streaming.
☆16Jul 28, 2026Updated 3 weeks ago
Alternatives and similar repositories for aws-glue-streaming-ingestion-from-kafka-to-apache-iceberg
Users that are interested in aws-glue-streaming-ingestion-from-kafka-to-apache-iceberg are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Streaming ETL job cases in AWS Glue to integrate Iceberg and creating an in-place updatable data lake on Amazon S3☆27Jul 28, 2026Updated 3 weeks ago
- ☆17Jan 11, 2024Updated 2 years ago
- ☆23Feb 14, 2025Updated last year
- 小売業で予測ベースの発注を実現するためのサンプルソリューション☆17Apr 10, 2025Updated last year
- This repository provides the resources required for the Amazon Redshift Streaming workshop☆13Apr 13, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This is a sample app created to illustrate the best practices outlined in the following blog post☆15Jul 28, 2026Updated 3 weeks ago
- ☆10Apr 5, 2024Updated 2 years ago
- ☆11May 7, 2024Updated 2 years ago
- ☆15Feb 12, 2026Updated 6 months ago
- This repository contains an analysis of the effects of COVID-19 on trade trends up to December 2021. The dataset used provides daily trad…☆16Aug 16, 2023Updated 3 years ago
- Stream CDC into an Amazon S3 data lake in Apache Iceberg table format with AWS Glue Streaming and DMS☆36Jul 28, 2026Updated 3 weeks ago
- Building a Q&A app (powered by a LLM model) using AWS Bedrock, AWS Kendra, AWS S3 and Streamlit in just a couple of hours☆17Dec 7, 2023Updated 2 years ago
- ☆19Dec 23, 2022Updated 3 years ago
- ☆21Nov 11, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This repository contains End-to-end Data Analytics Projects☆17Jun 6, 2023Updated 3 years ago
- Data Analysis of Bicycle Manufacturing Company Using Python, SQL and Power BI☆14Apr 14, 2023Updated 3 years ago
- #DataPipeLine #ETL - Created is a Facebook data extraction utility to extract the publicly available data on Facebook. Used Facebook Grap…☆13Jun 27, 2018Updated 8 years ago
- Building a GPT-4 Q&A app using Azure OpenAI, Pinecone and Streamlit in just a couple of hours☆23Jul 6, 2023Updated 3 years ago
- ☆15Feb 20, 2018Updated 8 years ago
- Collection of code examples for Amazon Managed Service for Apache Flink☆90Jun 16, 2026Updated 2 months ago
- ☆26Apr 26, 2026Updated 3 months ago
- In this repository, explore insightful solutions through exploratory data analysis focusing on mental health problems. Gain valuable insi…☆26Dec 18, 2023Updated 2 years ago
- Removes unused images from AWS ECR☆19Apr 7, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Amazon Bedrock AI Karaoke is an interactive demonstration of Amazon Bedrock. Users complete the prompt with the microphone and choose the…☆20Jan 29, 2025Updated last year
- Delta Lake helper methods. No Spark dependency.☆21Jan 19, 2026Updated 6 months ago
- ☆14Feb 26, 2024Updated 2 years ago
- Fundamentals of Apache Flink [video], published by Packt☆12Jan 30, 2023Updated 3 years ago
- CLI tool to help importing existing dbt Cloud config to Terraform