NEW: see http://www.hops.io/. OLD: This work aims to re-engineer the Hadoop Distributed File System (HDFS) so that it can be 1) highly available, and 2) horizontally scalable. This is achieved by replacing the central master server with a distributed real-time database (in our implementation, MySQL Cluster).
☆26Jan 2, 2012Updated 14 years ago
Alternatives and similar repositories for Scaling-HDFS-NameNode
Users that are interested in Scaling-HDFS-NameNode are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A toy school project intended to be an approximate clone of Google's Megastore database for geographically-distributed scalable fault-to…☆35Oct 12, 2011Updated 14 years ago
- Bigtop is a project for the development of packaging and tests of the Apache Hadoop ecosystem. The primary goal of Bigtop is to build a …☆51Jul 4, 2011Updated 15 years ago
- An Object Storage System implementation based on Hadoop and HBase, with similar features like S3 (Amazon Simple Storage Service).☆19Apr 1, 2013Updated 13 years ago
- Patched, refactored version of code.google.com/hadoop-gpl-compression for hadoop 0.20☆37Aug 13, 2012Updated 14 years ago
- Codec for Hadoop adding OpenPGP encryption using Bouncy Castle☆17Aug 18, 2011Updated 15 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Patched version of Cloudera's Distribution of Hadoop with Mesos support☆13Nov 6, 2011Updated 14 years ago
- A weather monitoring Dashboard built upon Python and Yahoo API☆14Jun 10, 2015Updated 11 years ago
- Find implementation for Hadoop☆17Sep 9, 2015Updated 11 years ago
- A simple benchmark of noSQL databases for both read/update and MapReduce performances☆32May 14, 2011Updated 15 years ago
- Continuous Streaming SQL Queries for Flume☆96Dec 30, 2011Updated 14 years ago
- Framework that makes processing arbitrary binary data in Hadoop easier☆22Apr 8, 2013Updated 13 years ago
- Demo re-implementation of the Hadoop MapReduce scheduler in Python☆13Mar 1, 2016Updated 10 years ago
- Utilities for HBase cluster management☆15May 2, 2012Updated 14 years ago
- JVMTI agent which calls mlockall and setuids down to a target user upon initialization☆21Sep 13, 2011Updated 15 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Provides canary based prewarming of lambda functions for Kinesis Event Sources.☆15Oct 13, 2020Updated 5 years ago
- ALPS: An Adaptive Learning, Priority OS Scheduler for Serverless Functions (USENIX ATC'24)☆15Jun 20, 2024Updated 2 years ago
- S3 like FileSystem based on Cassandra☆34May 14, 2010Updated 16 years ago
- λFS: an elastic, high-performance, serverless-function-based metadata service for large-scale distributed file systems (ACM ASPLOS'23)☆14Apr 2, 2025Updated last year
- Kafka, Spark Streaming, Kudu integration examples☆17Dec 22, 2017Updated 8 years ago
- BESPOKV: Application-Tailored Flexible Key-Value Store for HPC☆12Aug 28, 2018Updated 8 years ago
- A big data cluster management tool that creates and manages clusters of different technologies.☆21Apr 20, 2015Updated 11 years ago
- FLY a Domain Specific Language for scientific computing on the Multi Cloud☆12Apr 11, 2023Updated 3 years ago
- A Hanborq optimized Hadoop Distribution, especially with high performance of MapReduce. It's the core part of HDH (Hanborq Distribution w…☆50Mar 26, 2012Updated 14 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆16May 11, 2022Updated 4 years ago
- a column file format☆133Sep 25, 2012Updated 14 years ago
- Hadoop Data Integration with various databases, ftp servers, salesforce. Incremental update, dedup, append, merge your data on Hadoop.☆92Apr 11, 2013Updated 13 years ago
- HBase as the backing store for the TF-IDF representations for Lucene☆110May 14, 2010Updated 16 years ago
- Library to call AWS Lambda functions preventing duplicate execution☆18Jun 14, 2015Updated 11 years ago
- Mahout vector encoding for pig☆53Nov 20, 2022Updated 3 years ago
- Implementation of Tyler Neylon's Locality-Specific Hash based on simplex tesselations☆28Oct 15, 2011Updated 14 years ago
- Albis: High-Performance File Format for Big Data Systems☆21Jul 12, 2018Updated 8 years ago
- a UI for OpenTSDB☆138May 14, 2014Updated 12 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Integration code to enable Hadoop processing of data in NetCDF format☆31May 22, 2013Updated 13 years ago
- Simple fault-tolerant distributed file system. 1. Utilize virtual ring style key-value store with 3 backups. 2. Automated failure detecti…☆16Jan 22, 2016Updated 10 years ago
- Forward webhook to multiple destinations using Google Cloud Functions☆11May 9, 2023Updated 3 years ago
- Whether you're running a call center, just a small office, or have to deal with all of the incoming calls for that event tonight, puttin…☆17Mar 19, 2015Updated 11 years ago
- ☆27Apr 17, 2019Updated 7 years ago
- Benchmarking for custom Redis commands and modules☆33Jan 25, 2021Updated 5 years ago
- Tuple MapReduce for Hadoop: Hadoop API made easy☆57Jun 27, 2022Updated 4 years ago