Timberlake is a Job Tracker for Hadoop.
☆177Jan 24, 2020Updated 6 years ago
Alternatives and similar repositories for timberlake
Users that are interested in timberlake are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tools for working with parquet, impala, and hive☆135Jan 4, 2021Updated 5 years ago
- A key/value store for serving static batch data☆174Jul 14, 2023Updated 3 years ago
- ☆29Nov 17, 2014Updated 11 years ago
- scalding powered machine learning☆109Nov 18, 2014Updated 11 years ago
- Reduce your data. A unix filter for algebird-powered aggregation.☆141Apr 17, 2017Updated 9 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Collect local Mesos slave, underlying operating system and machine metrics and produce to Apache Kafka☆20Jan 29, 2016Updated 10 years ago
- hRaven collects run time data and statistics from MapReduce jobs in an easily queryable format☆129Jan 14, 2022Updated 4 years ago
- A Hadoop map reduce framework for Scala.☆15Apr 21, 2016Updated 10 years ago
- Useful reusable pipeline components for Crunch jobs☆27Feb 10, 2015Updated 11 years ago
- Storehaus is a library that makes it easy to work with asynchronous key value stores☆465Jul 17, 2020Updated 6 years ago
- On demand presto cluster with mesos, marathon and docker.☆29Mar 7, 2018Updated 8 years ago
- Deploying apache-hadoop in a virtualized cluster as easy as 1-2-3.☆15Jul 17, 2019Updated 7 years ago
- Scripts for parsing / making sense of yarn logs☆51Aug 22, 2016Updated 9 years ago
- Memcached on YARN☆19Jun 2, 2014Updated 12 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Quick up and running using Go for Apache Kafka☆22Oct 4, 2015Updated 10 years ago
- Verspätung is a small utility which aims to help identify delay of Kafka consumers☆11Aug 4, 2016Updated 9 years ago
- This repository implements converters and tools for working with NGS data in HPC or Hadoop cluster☆17Apr 13, 2018Updated 8 years ago
- an impala client for ruby☆34Jan 25, 2017Updated 9 years ago
- Next-generation web analytics processing with Scala, Spark, and Parquet.☆330Mar 28, 2015Updated 11 years ago
- Visualize your HDFS cluster usage☆228Oct 13, 2020Updated 5 years ago
- Streaming MapReduce with Scalding and Storm☆2,123Jan 19, 2022Updated 4 years ago
- Protobuf support for Finagle☆14Nov 7, 2022Updated 3 years ago
- Parallel Genomic Analysis Toolkit☆14Feb 11, 2019Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Hadoop Tool implementation which enables extreme productivity - running MR jobs on your cluster right from your sbt shell!☆19Feb 2, 2014Updated 12 years ago
- Serving system for batch generated data sets☆179May 11, 2017Updated 9 years ago
- Summingbird Workshop at Lambda Jam 2013.☆24Aug 21, 2018Updated 7 years ago
- source examples to support the "Cascading for the Impatient" blog post series☆79Aug 30, 2016Updated 9 years ago
- The code for the in memory data pipeline that was presented at Berlin Buzzwords 2015.☆10Jun 1, 2015Updated 11 years ago
- Cassandra state implementation for Twitter Storm Trident API☆17Jan 21, 2013Updated 13 years ago
- Prescriptive Applications over Kite and Hadoop☆12Oct 14, 2015Updated 10 years ago
- A tool for running Spark on Google Compute Engine☆16Jan 20, 2017Updated 9 years ago
- Combination of Dockerized Hortonworks projects and other Hadoop ecosystem components☆10Oct 11, 2019Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A platform for visualization and real-time monitoring of data workflows☆1,170Jan 22, 2020Updated 6 years ago
- ☆186Jun 12, 2015Updated 11 years ago
- ☆20Jul 15, 2015Updated 11 years ago
- An Apache Mesos Framework that allows for replaying load over and over and over (and over) again☆10Aug 10, 2015Updated 10 years ago
- Functional, Typesafe, Declarative Data Pipelines☆140Jan 29, 2018Updated 8 years ago
- A Scala API for Cascading☆3,522May 28, 2023Updated 3 years ago
- REST job server for Spark. Note that this is *not* the mainline open source version. For that, go to https://github.com/spark-jobserver…☆345May 19, 2017Updated 9 years ago