☆15Dec 14, 2010Updated 15 years ago
Alternatives and similar repositories for Flume-Hive
Users that are interested in Flume-Hive are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Extracts A Social Network From Cassandra NoSQL Data-store To The InfiniteGraph Graph Database For Analysis☆16Aug 26, 2010Updated 16 years ago
- A program to use MapReduce and Graph Theory to efficiently and scalably find all words in a Boggle roll.☆16Sep 5, 2013Updated 12 years ago
- Optimized joins using bloom filters on Hadoop via Cascading.☆22Sep 25, 2009Updated 16 years ago
- Text clustering service for the web☆25Mar 30, 2019Updated 7 years ago
- Simple Java Beans mapping for HBase☆24Jul 11, 2012Updated 14 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- HBase as the backing store for the TF-IDF representations for Lucene☆110May 14, 2010Updated 16 years ago
- ☆24Jan 19, 2012Updated 14 years ago
- S3 log bucket parser app for Django☆15Sep 12, 2011Updated 14 years ago
- A bunch of utility classes for Java, Hadoop, HBase, Pig, etc.☆77Mar 31, 2014Updated 12 years ago
- A deterministic list cache for Django☆15Aug 22, 2011Updated 15 years ago
- Mysql Puppet Module☆16Aug 12, 2016Updated 10 years ago
- A tool for loading data into elastic-search in bulk☆16Jun 20, 2013Updated 13 years ago
- A REST API for Mozilla Metrics services.☆59Mar 30, 2019Updated 7 years ago
- An example of using Hadoop to Cassandra through the Binary Memtable☆16Aug 25, 2009Updated 17 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Real-time Monitoring☆29May 14, 2012Updated 14 years ago
- Bixo is an open source web mining toolkit that runs as a series of Cascading pipes on top of Hadoop. By building a customized Cascading p…☆143Jul 7, 2022Updated 4 years ago
- Guide to Recommender Systems☆14Feb 24, 2012Updated 14 years ago
- Write tweets from the Twitter streaming API to Hadoop☆15Feb 20, 2010Updated 16 years ago
- SmartSocket is an extensible open source, Java and PHP socket server engine. Its aim is to make creating multi-user applications as quic…☆15May 26, 2012Updated 14 years ago
- Neo4j POC to Integrate VisualSearch.js and Cypher☆18May 31, 2016Updated 10 years ago
- Design of a specification for the automation of infrastructure deployments☆24Apr 6, 2022Updated 4 years ago
- Zohmg is a data store for aggregation of multi-dimensional time series data, built on top of Hadoop, Dumbo and HBase.☆173Oct 16, 2012Updated 13 years ago
- HBase an open-source implementation of Google’s Bigtable, a massively distributed, scalable, reliable, non-relational database.☆21May 18, 2017Updated 9 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Indexed HBase. An extnestion of HBASE core which support faster scans at the expense of larger RAM consumption.☆40Jul 8, 2010Updated 16 years ago
- `gen_cluster` is an erlang behavior for pid clustering. It is a cascading behavior that builds on `gen_server`.☆16May 6, 2010Updated 16 years ago
- Neo4j and the Crunchbase API mashup☆19Aug 15, 2013Updated 13 years ago
- HDFS endpoint collecting and aggregating data flows☆19Oct 24, 2013Updated 12 years ago
- This is a prototype app that store items into a Hazelcast map and queue based on the description in https://wiki.mozilla.org/Socorro:Clie…☆17Apr 11, 2011Updated 15 years ago
- A disk-based HashMap implementation allowing persistence of data across sessions.☆15May 7, 2014Updated 12 years ago
- Apache Camel component for Beanstalk☆16Sep 21, 2014Updated 11 years ago
- code etc for Think Stats book http://greenteapress.com/thinkstats/☆18Sep 22, 2011Updated 14 years ago
- thrudb - document oriented database services☆36Mar 16, 2010Updated 16 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A simple benchmark of noSQL databases for both read/update and MapReduce performances☆32May 14, 2011Updated 15 years ago
- Java classes that can be useful for Dumbo programs that run on Hadoop Streaming.☆26May 20, 2012Updated 14 years ago
- Distributed Java Collections for ZooKeeper☆112Jul 4, 2016Updated 10 years ago
- A script for configuring apache with Django+wsgi/fcgi+virtualenv+etc.☆22Nov 19, 2009Updated 16 years ago
- Slides and Examples From "DTrace and Erlang" @ BashoChat 001☆16Dec 15, 2011Updated 14 years ago
- Make your Hibernate Search more Elastic ! WARNING : project suspended !☆16Apr 24, 2011Updated 15 years ago
- Script backups easily to S3 using Python☆17Feb 1, 2022Updated 4 years ago