Launch Spark clusters on YARN
☆24Aug 29, 2011Updated 15 years ago
Alternatives and similar repositories for spark-yarn
Users that are interested in spark-yarn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Spooker is a dynamic framework for processing high volume data streams via processing pipelines☆30Feb 1, 2016Updated 10 years ago
- Storm-yarn enables Storm clusters to be deployed into machines managed by Hadoop YARN.☆415Jul 21, 2023Updated 3 years ago
- Friendly ML feature store☆45May 19, 2022Updated 4 years ago
- Parallel Algorithms in Python for Hadoop/Mapreduce☆55Aug 10, 2012Updated 14 years ago
- An example of using Flink for Fault-Tolerant Stream Processing☆12Jan 3, 2019Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A collection of python tools for extracting descriptors from images (whole and sub-image descriptors).☆17Dec 4, 2017Updated 8 years ago
- A Load UDF for loading JSON files with Pig☆15Jul 6, 2011Updated 15 years ago
- Generic implementation of Information Theory-based Feature Selection methods. It also contains an Entropy Minimization Discretization imp…☆19Jul 21, 2014Updated 12 years ago
- SipHash implementation with hand inlining the SIPROUND☆15Jun 8, 2014Updated 12 years ago
- Project superseded by https://github.com/mesos/storm☆25Sep 30, 2014Updated 11 years ago
- A reporistory of User-defined functions for Apache Pig☆16Sep 20, 2010Updated 15 years ago
- Main Repo☆15Jun 24, 2010Updated 16 years ago
- akka 中文文档,通过官方文档翻译☆27Mar 16, 2015Updated 11 years ago
- Lua-MapReduce framework implemented in Lua using luamongo driver and MongoDB as storage. It follows Iterative MapReduce for training of M…☆25Dec 23, 2015Updated 10 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Bayesian data analysis with PyMC(2)☆17Oct 11, 2013Updated 12 years ago
- Library and a Framework for building fast, scalable, fault-tolerant Data APIs based on Akka, Avro, ZooKeeper and Kafka☆25Oct 16, 2020Updated 5 years ago
- Low level integration of Spark and Kafka☆129Mar 15, 2018Updated 8 years ago
- Trident Function to be able to call R functions within a Trident Topology☆47Dec 16, 2023Updated 2 years ago
- 最快速精简的java消息队列,可以比较容易的基于这个项目进行二次开发☆32Dec 14, 2016Updated 9 years ago
- HBase as a JSON Document Database☆27Jun 14, 2023Updated 3 years ago
- Meta-repository of big data tools -- source and essential plugins for hadoop, pig, wukong, storm, kafka etc.☆30Jun 29, 2014Updated 12 years ago
- ☆17Jan 7, 2026Updated 8 months ago
- Functions for parsing INI-style file natively in bash☆18Jun 4, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Talks and demonstration notebooks about IPython, to be used by anyone (all materials CC-BY licensed)☆17Oct 31, 2012Updated 13 years ago
- ☆10Aug 28, 2014Updated 12 years ago
- 关于Spark的源码分析,以及平时工作的一些总结☆31Dec 25, 2015Updated 10 years ago
- C++ APIs for Alluxio (formerly Tachyon)☆19Nov 29, 2016Updated 9 years ago
- Files for the intro to D3 workshop☆17Oct 21, 2013Updated 12 years ago
- A streaming key-value store implementation using native Flink Streaming operators☆22Oct 10, 2015Updated 10 years ago
- Serving system for batch generated data sets☆179May 11, 2017Updated 9 years ago
- Fast, reliable, and scalable channels implementation based on Redis streams.☆11Jun 25, 2024Updated 2 years ago
- Hue based HBase Explorer☆25Dec 14, 2010Updated 15 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Experiments with VowPal Wabbit Machine Learning & Storm☆26Apr 29, 2013Updated 13 years ago
- SNA project☆26Nov 11, 2012Updated 13 years ago
- Zephyr is a big data, platform agnostic ETL API, with Hadoop MapReduce, Storm, and other big data bindings.☆22Jan 21, 2016Updated 10 years ago
- VSCode extension for SSH config files.☆13Jul 31, 2019Updated 7 years ago
- The fast and fun way to write YARN applications.☆135Nov 14, 2018Updated 7 years ago
- Materials related to the PyData workshop / PyCon 2012 tutorials☆19Mar 21, 2012Updated 14 years ago
- ☆15Sep 29, 2015Updated 10 years ago