Maelstrom is an open source Kafka integration with Spark that is designed to be developer friendly, high performance (millisecond stream processing), scalable (consumes messges at Spark worker nodes), and is extremely reliable.
☆21Feb 6, 2017Updated 9 years ago
Alternatives and similar repositories for maelstrom
Users that are interested in maelstrom are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Distributed SQL base Realtime Streaming Computation Framework On Apache Storm, Spark☆12Mar 14, 2016Updated 10 years ago
- Library and a Framework for building fast, scalable, fault-tolerant Data APIs based on Akka, Avro, ZooKeeper and Kafka☆25Oct 16, 2020Updated 5 years ago
- Open source task scheduler with dependency management☆15Jul 1, 2018Updated 8 years ago
- A quotation-based Scala DSL for scalable data analysis.☆65Jul 7, 2022Updated 4 years ago
- Spark CEP is an extension of Spark Streaming to support SQL-based query processing☆58Apr 12, 2017Updated 9 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Open-source distribute workflow schedule tools, also support streaming task.☆40Nov 11, 2017Updated 8 years ago
- DataStax Enterprise (DSE) Deployment Guide for Google Cloud Platform (GCP)☆10Apr 10, 2020Updated 6 years ago
- Scala API for Apache Spark SQL high-order functions☆15Aug 4, 2023Updated 3 years ago
- Temporal_Graph_library☆25Feb 2, 2019Updated 7 years ago
- A set of tools to ease working with Zookeeper and Kafka.☆23Jan 22, 2016Updated 10 years ago
- Plot live-stats as graph from ApacheSpark application using Lightning-viz☆18Jul 3, 2017Updated 9 years ago
- A distributed generic query layer for Apache Kafka Interactive Queries☆26Nov 8, 2017Updated 8 years ago
- Demonstration of a Hive Input Format for Iceberg☆26Mar 12, 2021Updated 5 years ago
- A core AST and utilities to manipulate geographical data☆21Sep 30, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Euphoria is an open source Java API for creating unified big-data processing flows. It provides an engine independent programming model w…☆81Nov 15, 2022Updated 3 years ago
- Parquet file generator☆22Apr 17, 2018Updated 8 years ago
- Cassandra state implementation for Twitter Storm Trident API☆17Jan 21, 2013Updated 13 years ago
- ☆51Jul 24, 2026Updated 2 weeks ago
- Presto connector for Apache Kudu☆48Mar 22, 2019Updated 7 years ago
- Protobuf serialization support for Apache Flink☆22Jun 1, 2021Updated 5 years ago
- an open source dataworks platform☆20Jun 4, 2021Updated 5 years ago
- Framework for running macro benchmarks in a clustered environment☆25Aug 29, 2022Updated 3 years ago
- Monitoring cassandra cluster by ELK (Elasticsearch , logstash and Kibana)☆20Mar 16, 2017Updated 9 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- UberScriptQuery, a SQL-like DSL to make writing Spark jobs super easy☆65Dec 17, 2023Updated 2 years ago
- Flink performance tests☆29Nov 13, 2019Updated 6 years ago
- Example code for building your own MemSQL Streamliner Pipelines☆23Apr 18, 2017Updated 9 years ago
- Serverless proxy for Spark cluster☆326Apr 13, 2026Updated 3 months ago
- akkaflow是一个基于akka架构上构建的分布式高可用DAG工作流调度工具,可以把子节点分配在集群机器上并行执行,高效利用集群资源。☆106Sep 14, 2019Updated 6 years ago
- Zephyr is a big data, platform agnostic ETL API, with Hadoop MapReduce, Storm, and other big data bindings.☆21Jan 21, 2016Updated 10 years ago
- Web Based Kafka Consumer and Producer☆70Jan 29, 2020Updated 6 years ago
- Fast, reliable, and scalable channels implementation based on Redis streams.☆11Jun 25, 2024Updated 2 years ago
- li-apache-kafka-clients is a wrapper library for the Apache Kafka vanilla clients. It provides additional features such as large message …☆137Jul 7, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Apache Geode on Kubernetes☆10Oct 19, 2019Updated 6 years ago
- A simple golang job queue☆13Jan 19, 2023Updated 3 years ago
- Write your Spark data to Kafka seamlessly☆172Jul 10, 2024Updated 2 years ago
- kubectl plugin for generating nginx-ingress compatible basic-auth secrets on kubernetes clusters☆21Jan 9, 2022Updated 4 years ago
- Provides an abstract data access layer on top of metric stores. Supports both SQL and structured JSON queries.☆22Nov 6, 2015Updated 10 years ago
- Detect duplicated items。内容排重框架。☆11Apr 30, 2015Updated 11 years ago
- A light Kafka to HDFS/S3 ETL library based on Apache Spark☆40Jun 29, 2017Updated 9 years ago