Kafka stream for Spark with storage of the offsets in ZooKeeper
☆60Apr 18, 2017Updated 9 years ago
Alternatives and similar repositories for spark-kafka-source
Users that are interested in spark-kafka-source are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for processing AVRO data in Spark Streaming + Kafka (DirectKafka approach with custom offset management in zookeeper)☆29Sep 9, 2016Updated 9 years ago
- ☆241Jun 14, 2018Updated 8 years ago
- 使用shell脚本部署Apache Doris (incubating) FE & BE☆11Jul 8, 2019Updated 7 years ago
- ☆48Feb 4, 2018Updated 8 years ago
- spark实例代码☆78Nov 11, 2017Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Ingress data from kafka topic into clickhouse table (JSON format)☆24Apr 12, 2018Updated 8 years ago
- SparkStreaming中利用MySQL保存Kafka偏移量保证0数据丢失☆43Aug 2, 2017Updated 9 years ago
- DirectKafka examples for Spark Streaming : 1. with checkpointing 2. Custom offset management☆59Sep 9, 2016Updated 9 years ago
- High Performance Kafka Connector for Spark Streaming.Supports Multi Topic Fetch, Kafka Security. Reliable offset management in Zookeeper.…☆631Apr 24, 2026Updated 3 months ago
- Write your Spark data to Kafka seamlessly☆172Jul 10, 2024Updated 2 years ago
- 基于TBSchedule开发的一个分布式任务调度框架,可以解析任务间的依赖,并执行任务(执行Shell、bat脚本)☆12Aug 5, 2016Updated 10 years ago
- Clickhouse typesafe RowBinary insert tooling☆13Jul 6, 2019Updated 7 years ago
- AWS SSM in Action, the next generation of SSH☆23Mar 14, 2018Updated 8 years ago
- Example usage of spark cassandra connector☆25Nov 21, 2014Updated 11 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Ansible playbooks to help to deploy Apache Hadoop,Spark,Storm,Zookeeper,Elasticsearch,Azkaban,Flume,Hbase,Kafka,Kibana,Logstash☆10Mar 21, 2017Updated 9 years ago
- 基于 spark 混合查询平台,支持不同源数据库的联合查询,mysql hive presto ...☆14Aug 3, 2017Updated 9 years ago
- Spark(multi versions) + Streaming/Hive/SQL/UDF Demos☆15May 17, 2018Updated 8 years ago
- A Kafka metric sink for Apache Spark☆11Apr 13, 2017Updated 9 years ago
- A Real-Time Analytical Processing (RTAP) example using Spark/Shark☆51Feb 21, 2014Updated 12 years ago
- Real Time Streaming using Apache Spark Streaming [Video], published by Packt☆10Oct 31, 2022Updated 3 years ago
- Pinot 是一个实时分布式的 OLAP 数据存储和分析系统。LinkedIn 使用它实现低延迟可伸缩的实时分析。Pinot 从离线数据源(包括 Hadoop 和各类文件)和在线数据源(如 Kafka)中攫取数据进行分析。Pinot 被设计是可以进行水平扩展的☆16Nov 8, 2015Updated 10 years ago
- Notes about Spark Streaming in Apache Spark☆59Apr 7, 2017Updated 9 years ago
- 记录Spark、Flink研究经验☆25Aug 11, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A testing DSL for kafka-streams☆14Aug 9, 2017Updated 9 years ago
- Java调用Kettle API执行转换和作业,Java代码生成Kettle转换。☆21Mar 9, 2018Updated 8 years ago
- Hands-On Data Warehousing with Azure Data Factory, published by Packt☆15Jan 18, 2023Updated 3 years ago
- HBase RDD example project☆19Jan 22, 2021Updated 5 years ago
- Store batched Kafka messages in S3.☆39Apr 13, 2022Updated 4 years ago
- An introduction of Scala learning and some frequently asked questions(FAQ);有关Scala的学习笔记,记录Scala的常用语法及标准库的部分设计原理☆26May 12, 2016Updated 10 years ago
- NuCypher for Kafka. Start building from this module (it fetches the appropriate branch from Kafka repository)☆17Oct 13, 2017Updated 8 years ago
- Metabase Impala Driver☆11May 28, 2024Updated 2 years ago
- A simple Spark LDA example. to demonstrate a full fletched clustering algorithm, with data cleaning using the processess like lemmatizati…☆23Oct 8, 2016Updated 9 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A type driven approach to string interpolation, aiming at consistent, secure, and only-human-readable logs and console outputs !☆14May 5, 2024Updated 2 years ago
- This project is a unified ETL platform that support various data processing technologies, including Spark, Hive, Hadoop, Python, Linux Sh…☆17Oct 16, 2015Updated 10 years ago
- This application comes as Spark2.1-as-Service-Provider using an embedded, Reactive-Streams-based, fully asynchronous HTTP server☆50Jul 16, 2023Updated 3 years ago
- Project for reading data from kafka and writing to kafka and HBase with kerberos☆24Dec 8, 2016Updated 9 years ago
- Some extra directives and other things for akka-http projects☆32Oct 29, 2017Updated 8 years ago
- Spark structured streaming with Kafka data source and writing to Cassandra☆62Dec 5, 2019Updated 6 years ago
- something to help you spark☆19Jul 5, 2017Updated 9 years ago