Timberlake is a Job Tracker for Hadoop.
☆177Jan 24, 2020Updated 6 years ago
Alternatives and similar repositories for timberlake
Users that are interested in timberlake are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tools for working with parquet, impala, and hive☆135Jan 4, 2021Updated 5 years ago
- A key/value store for serving static batch data☆174Jul 14, 2023Updated 3 years ago
- ☆29Nov 17, 2014Updated 11 years ago
- scalding powered machine learning☆109Nov 18, 2014Updated 11 years ago
- Reduce your data. A unix filter for algebird-powered aggregation.☆140Apr 17, 2017Updated 9 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Collect local Mesos slave, underlying operating system and machine metrics and produce to Apache Kafka☆20Jan 29, 2016Updated 10 years ago
- hRaven collects run time data and statistics from MapReduce jobs in an easily queryable format☆129Jan 14, 2022Updated 4 years ago
- A Hadoop map reduce framework for Scala.☆15Apr 21, 2016Updated 10 years ago
- Useful reusable pipeline components for Crunch jobs☆27Feb 10, 2015Updated 11 years ago
- Storehaus is a library that makes it easy to work with asynchronous key value stores☆465Jul 17, 2020Updated 6 years ago
- On demand presto cluster with mesos, marathon and docker.☆29Mar 7, 2018Updated 8 years ago
- Deploying apache-hadoop in a virtualized cluster as easy as 1-2-3.☆15Jul 17, 2019Updated 7 years ago
- Scripts for parsing / making sense of yarn logs☆51Aug 22, 2016Updated 9 years ago
- Quick up and running using Go for Apache Kafka☆22Oct 4, 2015Updated 10 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Memcached on YARN☆19Jun 2, 2014Updated 12 years ago
- Verspätung is a small utility which aims to help identify delay of Kafka consumers☆11Aug 4, 2016Updated 10 years ago
- Schedoscope is a scheduling framework for painfree agile development, testing, (re)loading, and monitoring of your datahub, lake, or what…☆98Nov 14, 2019Updated 6 years ago
- Static analysis tool for Golang that protects against SQL injections☆558Sep 20, 2021Updated 4 years ago
- This repository implements converters and tools for working with NGS data in HPC or Hadoop cluster☆17Apr 13, 2018Updated 8 years ago
- Visualize your HDFS cluster usage☆228Oct 13, 2020Updated 5 years ago
- Next-generation web analytics processing with Scala, Spark, and Parquet.☆330Mar 28, 2015Updated 11 years ago
- Streaming MapReduce with Scalding and Storm☆2,123Jan 19, 2022Updated 4 years ago
- Protobuf support for Finagle☆14Nov 7, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Parallel Genomic Analysis Toolkit☆14Feb 11, 2019Updated 7 years ago
- Hadoop Tool implementation which enables extreme productivity - running MR jobs on your cluster right from your sbt shell!☆19Feb 2, 2014Updated 12 years ago
- Serving system for batch generated data sets☆179May 11, 2017Updated 9 years ago
- Summingbird Workshop at Lambda Jam 2013.☆24Aug 21, 2018Updated 7 years ago
- Scala stuff☆18Jun 13, 2019Updated 7 years ago
- source examples to support the "Cascading for the Impatient" blog post series☆79Aug 30, 2016Updated 9 years ago
- Cassandra state implementation for Twitter Storm Trident API☆17Jan 21, 2013Updated 13 years ago
- Prescriptive Applications over Kite and Hadoop☆12Oct 14, 2015Updated 10 years ago
- The code for the in memory data pipeline that was presented at Berlin Buzzwords 2015.☆10Jun 1, 2015Updated 11 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A tool for running Spark on Google Compute Engine☆16Jan 20, 2017Updated 9 years ago
- Combination of Dockerized Hortonworks projects and other Hadoop ecosystem components☆10Oct 11, 2019Updated 6 years ago
- A platform for visualization and real-time monitoring of data workflows☆1,170Jan 22, 2020Updated 6 years ago
- ☆186Jun 12, 2015Updated 11 years ago
- An Apache Mesos Framework that allows for replaying load over and over and over (and over) again☆10Aug 10, 2015Updated 11 years ago
- ☆20Jul 15, 2015Updated 11 years ago
- Functional, Typesafe, Declarative Data Pipelines☆140Jan 29, 2018Updated 8 years ago