My data is bigger than your data!
☆39Aug 21, 2026Updated last week
Alternatives and similar repositories for my-data-is-bigger-than-your-data
Users that are interested in my-data-is-bigger-than-your-data are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Benchmarking Java fibers for IO usage☆12May 25, 2020Updated 6 years ago
- Example using Grafana with Druid☆11Mar 27, 2015Updated 11 years ago
- Emacs configuration files.☆11Jul 28, 2023Updated 3 years ago
- Begin here to create an AngularJS-based application.☆15Sep 25, 2014Updated 11 years ago
- HDFS Automatic Snapshot Service for Linux☆11Oct 17, 2016Updated 9 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An expansive bundle of NiFi additions intended to be used for generating test data☆11Aug 6, 2023Updated 3 years ago
- Mac OS X VNC client☆15May 3, 2020Updated 6 years ago
- Capstan example project for Java applications☆20Sep 4, 2016Updated 9 years ago
- Tool to squash the current branch/HEAD into a single rebased commit☆14Apr 27, 2017Updated 9 years ago
- Hadoop log aggregator and dashboard☆190Oct 29, 2013Updated 12 years ago
- Tweet Analysis with Spark☆13Aug 28, 2017Updated 9 years ago
- Mirror of Apache Pig☆18Jul 9, 2013Updated 13 years ago
- Dotfiles that I use (in bash)☆17Updated this week
- Reveal.js version of my "Keeping WordPress Under [Version] Control with Git" blog post☆17Jun 8, 2017Updated 9 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Tiny benchmarking framework for Java 7+☆17Jan 10, 2017Updated 9 years ago
- Tweet probabilistically generated HN post titles.☆29Nov 27, 2022Updated 3 years ago
- Use Vagrant to manage Rackspace Cloud instances.☆24Jan 15, 2016Updated 10 years ago
- sitemap.xml generation using lxml with support for alternates.☆13Aug 24, 2026Updated last week
- Template library for multiple formats, which can interoperate☆16Jan 25, 2015Updated 11 years ago
- A unit testing framework for the Cascading data processing platform.☆24Aug 25, 2021Updated 5 years ago
- Distributed version restore tool for S3☆12Jan 5, 2015Updated 11 years ago
- Boilerplate project for MOTW Workshop 2015☆10Mar 3, 2016Updated 10 years ago
- $HOME is where these files go; there are many like them, but these are mine...☆23Aug 15, 2026Updated 2 weeks ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Apache Avro™ support plugin for IntelliJ☆19Oct 5, 2017Updated 8 years ago
- Tuple MapReduce for Hadoop: Hadoop API made easy☆57Jun 27, 2022Updated 4 years ago
- Site-wide perimeter access control for Django projects☆17Mar 30, 2026Updated 5 months ago
- Helper to easily load fixtures in Django 1.7 data migrations.☆13Jul 24, 2015Updated 11 years ago
- Tools for working with parquet, impala, and hive☆135Jan 4, 2021Updated 5 years ago
- Django app providing a foreign key constraint support multiple fields☆11Jun 3, 2025Updated last year
- Rails app providing API to send audio for captioning using DeepSpeech☆18Apr 12, 2022Updated 4 years ago
- Android app for managing SickBeard.☆22Mar 28, 2018Updated 8 years ago
- A simple app that allows using Multilingual contentblock using Placeholders from django-cms☆14May 12, 2015Updated 11 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Installs Azul's Zulu OpenJDK 8 on a minimal Docker image☆55Jun 7, 2017Updated 9 years ago
- A CLI Plugin for Cloud Foundry for use with Brooklyn Service Broker☆10Jul 7, 2017Updated 9 years ago
- Blog running on jamescooke.info☆12May 17, 2024Updated 2 years ago
- A simple implementation of k-means clustering on the Spark cluster computing framework. See http://cs.berkeley.edu/~matei/spark.☆26Apr 9, 2011Updated 15 years ago
- ☆11Feb 15, 2023Updated 3 years ago
- JSON Schema files for Pandoc JSON☆14Aug 19, 2014Updated 12 years ago
- Zohmg is a data store for aggregation of multi-dimensional time series data, built on top of Hadoop, Dumbo and HBase.☆173Oct 16, 2012Updated 13 years ago