BlinkDB: Sub-Second Approximate Queries on Very Large Data.
☆660Feb 6, 2014Updated 12 years ago
Alternatives and similar repositories for blinkdb
Users that are interested in blinkdb are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Development in Shark has been ended.☆992Aug 11, 2015Updated 10 years ago
- Project SnappyData - memory optimized analytics database, based on Apache Spark™ and Apache Geode™. Stream, Transact, Analyze, Predict in…☆1,033Nov 21, 2022Updated 3 years ago
- Interactive-Speed Analytics: 200x Faster, 200x Fewer Cluster Resources, Approximate Query Processing☆252Jan 20, 2021Updated 5 years ago
- Lightning-fast cluster computing in Java, Scala and Python.☆1,419Apr 8, 2014Updated 12 years ago
- Distributed Prometheus time series database☆1,466Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Abstract Algebra for Scala☆2,298Nov 21, 2025Updated 8 months ago
- Streaming MapReduce with Scalding and Storm☆2,123Jan 19, 2022Updated 4 years ago
- Simplifying robust end-to-end machine learning on Apache Spark.☆473Apr 18, 2017Updated 9 years ago
- scalding powered machine learning☆109Nov 18, 2014Updated 11 years ago
- Former GraphX development repository. GraphX has been merged into Apache Spark; please submit pull requests there.☆359Dec 5, 2022Updated 3 years ago
- A Scala API for Cascading☆3,522May 28, 2023Updated 3 years ago
- Reduce your data. A unix filter for algebird-powered aggregation.☆140Apr 17, 2017Updated 9 years ago
- REST job server for Spark. Note that this is *not* the mainline open source version. For that, go to https://github.com/spark-jobserver…☆345May 19, 2017Updated 9 years ago
- MADlib has moved to Apache MADlib (incubating). Please send pull requests to the Apache repository.☆508Feb 9, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Stream summarizer and cardinality estimator.☆2,265Nov 28, 2019Updated 6 years ago
- Scala extensions for Storm☆132Jun 7, 2019Updated 7 years ago
- Distributed Graph Database☆5,224Oct 19, 2022Updated 3 years ago
- ☆92Nov 15, 2015Updated 10 years ago
- A platform for visualization and real-time monitoring of data workflows☆1,170Jan 22, 2020Updated 6 years ago
- Please visit https://github.com/h2oai/h2o-3 for latest H2O☆2,253Oct 24, 2024Updated last year
- Apache Heron (Incubating) is a realtime, distributed, fault-tolerant stream processing engine from Twitter☆3,628Mar 1, 2023Updated 3 years ago
- This code base is retained for historical interest only, please visit Apache Incubator Repo for latest one☆559Oct 5, 2022Updated 3 years ago
- Distributed and fault-tolerant realtime computation: stream processing, continuous computation, distributed RPC, and more☆8,765Aug 16, 2017Updated 8 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Fault tolerant job scheduler for Mesos which handles dependencies and ISO8601 based schedules☆4,375Jun 29, 2022Updated 4 years ago
- Scala extensions for the Kryo serialization library☆617Aug 19, 2024Updated last year
- An efficient updatable key-value store for Apache Spark☆255Mar 11, 2017Updated 9 years ago
- ☆110Apr 17, 2017Updated 9 years ago
- Stratosphere is now Apache Flink.☆201Dec 16, 2023Updated 2 years ago
- Mirror of Apache Apex core☆350Jun 7, 2021Updated 5 years ago
- Apache Pinot - A realtime distributed OLAP datastore☆6,122Updated this week
- Cassovary is a simple big graph processing library for the JVM☆1,052Oct 8, 2021Updated 4 years ago
- A Scala productivity framework for Hadoop.☆479Jul 1, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Sparrow scheduling platform (U.C. Berkeley).☆328Jul 25, 2020Updated 6 years ago
- Lightweight real-time big data streaming engine over Akka☆756Updated this week
- Hadoop library for large-scale data processing, now an Apache Incubator project☆581Jul 8, 2014Updated 12 years ago
- Trident-ML : A realtime online machine learning library☆383Dec 16, 2023Updated 2 years ago
- Storehaus is a library that makes it easy to work with asynchronous key value stores☆465Jul 17, 2020Updated 6 years ago
- An API for Distributed Machine Learning☆156Sep 22, 2016Updated 9 years ago
- A software library of stochastic streaming algorithms, a.k.a. sketches.☆959Aug 2, 2026Updated last week