S4 is a general-purpose, distributed, scalable, partially fault-tolerant, pluggable platform that allows programmers to easily develop applications for processing continuous unbounded streams of data.
☆233Mar 4, 2011Updated 15 years ago
Alternatives and similar repositories for core
Users that are interested in core are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- S4 Communication Layer☆38Jan 21, 2011Updated 15 years ago
- Example applications☆50Mar 8, 2011Updated 15 years ago
- A compiler and runtime for Google's Sawzall language, optimized for Hadoop☆41Apr 26, 2013Updated 13 years ago
- Explorations relative to cloning FlumeJava☆94Oct 13, 2020Updated 5 years ago
- Common metadata layer for Hadoop's Map Reduce, Pig, and Hive☆77Feb 17, 2011Updated 15 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A small personal project to learn Clojure by implementing some simple machine learning algorithms☆29Oct 12, 2009Updated 16 years ago
- WE HAVE MOVED to Apache Incubator. https://cwiki.apache.org/FLUME/ . Flume is a distributed, reliable, and available service for effici…☆943May 26, 2021Updated 5 years ago
- Continuous Streaming SQL Queries for Flume☆96Dec 30, 2011Updated 14 years ago
- Gremlins is a python framework for fault-testing distributed systems☆123May 12, 2014Updated 12 years ago
- PLEASE NOTE: Mesos is now hosted in Apache git! Get it using git clone https://git-wip-us.apache.org/repos/asf/mesos.git☆416Jan 22, 2018Updated 8 years ago
- Solandra = Solr + Cassandra☆881Mar 9, 2016Updated 10 years ago
- Annotations and Classes for managing and executing dependent processes☆39Apr 15, 2021Updated 5 years ago
- S4 repository☆142Nov 29, 2011Updated 14 years ago
- Toolkit of simple scripts useful for managing Hadoop☆16May 3, 2012Updated 14 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Distributed database specialized in exporting key/value data from Hadoop☆558Jun 27, 2014Updated 12 years ago
- Twitter's collection of LZO and Protocol Buffer-related Hadoop, Pig, Hive, and HBase code.☆1,134Apr 10, 2023Updated 3 years ago
- Lightning-fast cluster computing in Java, Scala and Python.☆1,419Apr 8, 2014Updated 12 years ago
- Dynamic HTTP Routing and Load Balancing via AMQP☆62Oct 24, 2009Updated 16 years ago
- HBase as the backing store for the TF-IDF representations for Lucene☆110May 14, 2010Updated 16 years ago
- Actors library for Clojure☆112Oct 26, 2011Updated 14 years ago
- Phoebus is a distributed framework for large scale graph processing written in Erlang.☆384Jan 15, 2012Updated 14 years ago
- A distributed task queue worker designed for throughput, parallelism, and clustering.☆239Jun 13, 2023Updated 3 years ago
- Distributed and fault-tolerant realtime computation: stream processing, continuous computation, distributed RPC, and more☆8,770Aug 16, 2017Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Use Avro to store all your values in HBase instead of regular columns☆76Dec 1, 2017Updated 8 years ago
- GoldenOrb is an open-source implementation of Pregel, Google's graph processing framework☆293Jun 29, 2022Updated 4 years ago
- A simple Python class for running multiple URL fetches in parallel☆40Jun 25, 2011Updated 15 years ago
- [Archived] A flexible sharding framework for creating eventually-consistent distributed datastores☆2,246Mar 16, 2017Updated 9 years ago
- Zohmg is a data store for aggregation of multi-dimensional time series data, built on top of Hadoop, Dumbo and HBase.☆173Oct 16, 2012Updated 13 years ago
- A toy school project intended to be an approximate clone of Google's Megastore database for geographically-distributed scalable fault-to…☆35Oct 12, 2011Updated 14 years ago
- Honu is a large scale data collection and processing pipeline☆84Feb 4, 2011Updated 15 years ago
- A Hadoop toolkit for web-scale information retrieval research☆87Dec 12, 2014Updated 11 years ago
- Extracts A Social Network From Cassandra NoSQL Data-store To The InfiniteGraph Graph Database For Analysis☆16Aug 26, 2010Updated 15 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [UNMAINTAINED] A minimalistic Ruby web framework for publishing Linked Data.☆34Feb 2, 2010Updated 16 years ago
- a column file format☆133Sep 25, 2012Updated 13 years ago
- A distributed publish/subscribe messaging service☆564Jun 10, 2023Updated 3 years ago
- Microblogging using RabbitMQ and ejabberd☆75Apr 5, 2009Updated 17 years ago
- A distributed, fault-tolerant graph database☆3,317Mar 16, 2017Updated 9 years ago
- Web based console for cassandra☆106Oct 1, 2012Updated 13 years ago
- An idiomatic scala library for interacting with cassandra☆56Sep 30, 2010Updated 15 years ago