Statistical Workload Injector for MapReduce - Project at UC Berkeley AMP Lab
☆128May 29, 2014Updated 12 years ago
Alternatives and similar repositories for SWIM
Users that are interested in SWIM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Cluster simulator with far memory☆12Apr 28, 2020Updated 6 years ago
- Borg cluster traces from Google☆1,090Updated this week
- Rain is a statistics-based workload generation toolkit that uses parameterized and empirical distributions to model the different classes…☆35Nov 2, 2016Updated 9 years ago
- Hadoop Data Integration with various databases, ftp servers, salesforce. Incremental update, dedup, append, merge your data on Hadoop.☆92Apr 11, 2013Updated 13 years ago
- REST based interface for PIG execution☆25Dec 13, 2021Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Hadoop log aggregator and dashboard☆190Oct 29, 2013Updated 12 years ago
- Madlib port for Cloudera Impala☆32Mar 30, 2017Updated 9 years ago
- ☆20Jul 7, 2017Updated 9 years ago
- Automatically exported from code.google.com/p/cluster-scheduler-simulator☆173Jun 3, 2022Updated 4 years ago
- USC Version of Hadoop that includes HDFS-RAID. Erasure codes like Locally Repairable Codes (aka Simple Regenerating Code), Reed Solomon C…☆71Jul 18, 2013Updated 13 years ago
- Real-time analytics in Apache Flume☆51Feb 2, 2016Updated 10 years ago
- Provides a simple archetype to create MapReduce jobs with Maven.☆24Dec 3, 2010Updated 15 years ago
- Code for Ernest☆33Jul 6, 2023Updated 3 years ago
- A monitoring tool for Hadoop cluster☆28Sep 12, 2012Updated 13 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Mahout vector encoding for pig☆53Nov 20, 2022Updated 3 years ago
- Facebook's Realtime Distributed FS based on Apache Hadoop 0.20-append☆874Oct 10, 2014Updated 11 years ago
- Secondary index on HBase☆18Oct 24, 2015Updated 10 years ago
- Main Repo☆15Jun 24, 2010Updated 16 years ago
- Iterative Reduce☆22Jun 3, 2014Updated 12 years ago
- Testbench for experimenting with Apache Hive at any data scale.☆64Jul 10, 2017Updated 9 years ago
- Fast, predictable data analytics based on (and API-compatible with) Apache Spark☆26Oct 28, 2017Updated 8 years ago
- Zoe: Container Analytics as a Service -- mirror of https://gitlab.eurecom.fr/zoe/main/☆50Sep 8, 2019Updated 6 years ago
- SDN Troubleshooting System☆29Jan 19, 2015Updated 11 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Load testing tools for Flume☆18Jun 22, 2012Updated 14 years ago
- Pig Visualization framework☆466Mar 24, 2023Updated 3 years ago
- CDAP Cube Dataset Guide☆12Aug 26, 2017Updated 8 years ago
- A set of Hadoop utilities to make working with Hadoop a little easier.☆26Feb 11, 2020Updated 6 years ago
- ☆29Nov 17, 2014Updated 11 years ago
- Allows you to see where(datanodes) that contain a file in HDFS☆17Mar 16, 2013Updated 13 years ago
- Cluster Far Mem, framework to execute single job and multi job experiments using fastswap☆21Jan 12, 2024Updated 2 years ago
- 基于Docker和Kubernetes的应用托管与部署平台☆12Dec 27, 2022Updated 3 years ago
- A simple easy to use Hadoop map reduce workflow engine☆18Mar 30, 2012Updated 14 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A compiler and runtime for Google's Sawzall language, optimized for Hadoop☆41Apr 26, 2013Updated 13 years ago
- Giraffa FileSystem (Slack: giraffa-fs.slack.com)☆18Mar 8, 2017Updated 9 years ago
- Workshop for Hadoop Operations Best Practices☆10Feb 24, 2015Updated 11 years ago
- A simple implementation of k-means clustering on the Spark cluster computing framework. See http://cs.berkeley.edu/~matei/spark.☆26Apr 9, 2011Updated 15 years ago
- Data Management + Feed Processing Platform over Hadoop☆27May 8, 2013Updated 13 years ago
- TPC-H Benchmark on Cloudera Impala☆19Apr 25, 2013Updated 13 years ago
- Hadoop task level native runtime☆27Jan 25, 2013Updated 13 years ago