A Python wrapper for Cascading
☆220Dec 30, 2019Updated 6 years ago
Alternatives and similar repositories for pycascading
Users that are interested in pycascading are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- All development now happens over here: https://github.com/cwensel/cascading. Cascading is a feature rich API for defining and executing c…☆332Nov 29, 2018Updated 7 years ago
- simple simhashing in hadoop with cascading☆33May 9, 2011Updated 15 years ago
- hRaven collects run time data and statistics from MapReduce jobs in an easily queryable format☆129Jan 14, 2022Updated 4 years ago
- A grouping of Apache Pig examples.☆65Oct 13, 2020Updated 5 years ago
- A reporistory of User-defined functions for Apache Pig☆16Sep 20, 2010Updated 15 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Twitter's collection of LZO and Protocol Buffer-related Hadoop, Pig, Hive, and HBase code.☆1,134Apr 10, 2023Updated 3 years ago
- A Scala API for Cascading☆3,522May 28, 2023Updated 3 years ago
- Python module that allows one to easily write and run Hadoop programs.☆1,030Jan 9, 2018Updated 8 years ago
- All the Cascading taps you need and love.☆39Mar 4, 2019Updated 7 years ago
- The Scalding tutorial as a standalone SBT project☆51Oct 16, 2017Updated 8 years ago
- Common metadata layer for Hadoop's Map Reduce, Pig, and Hive☆77Feb 17, 2011Updated 15 years ago
- Elephant Twin is a framework for creating indexes in Hadoop☆99Oct 12, 2020Updated 5 years ago
- A platform for visualization and real-time monitoring of data workflows☆1,170Jan 22, 2020Updated 6 years ago
- Materials for Cascalog workshop☆18Sep 17, 2011Updated 14 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Streaming MapReduce with Scalding and Storm☆2,123Jan 19, 2022Updated 4 years ago
- A C interface to watchman☆47Jun 7, 2019Updated 7 years ago
- Framework that makes processing arbitrary binary data in Hadoop easier☆22Apr 8, 2013Updated 13 years ago
- Lightning-fast cluster computing in Java, Scala and Python.☆1,419Apr 8, 2014Updated 12 years ago
- Hadoop library for large-scale data processing, now an Apache Incubator project☆581Jul 8, 2014Updated 12 years ago
- Machine Learning for Cascading☆85Jun 12, 2015Updated 11 years ago
- Cascading is a feature rich API for defining and executing complex and fault tolerant data processing flows locally or on a cluster.☆355Apr 8, 2025Updated last year
- A Python client for the HBase Avro interface☆50Feb 1, 2016Updated 10 years ago
- Python MapReduce library written in Cython. Visit us in #hadoopy on freenode. See the link below for documentation and tutorials.☆243Jan 8, 2016Updated 10 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Optimized joins using bloom filters on Hadoop via Cascading.☆22Sep 25, 2009Updated 16 years ago
- Data processing on Hadoop without the hassle.☆1,373May 18, 2023Updated 3 years ago
- scalding powered machine learning☆109Nov 18, 2014Updated 11 years ago
- Watches files and records, or triggers actions, when they change.☆10Aug 26, 2015Updated 10 years ago
- cascading_ext is a collection of tools built on top of the Cascading platform which make it easy to build, debug, and run simple and high…☆58Feb 25, 2026Updated 5 months ago
- Bulk loading for elastic search☆186Dec 16, 2023Updated 2 years ago
- Main Repo☆15Jun 24, 2010Updated 16 years ago
- cascading schemes and taps for JDBC☆27Jun 15, 2016Updated 10 years ago
- Hadoop Input and Ouput formats for MongoDB☆29Nov 15, 2011Updated 14 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A simple key/value store for indexing meta data on JSON-type fields☆46Apr 25, 2011Updated 15 years ago
- A simple standalone Finagle stats viewer☆22Aug 28, 2013Updated 12 years ago
- Scala extensions for Storm☆132Jun 7, 2019Updated 7 years ago
- Utilities for working with Hadoop and Cascading☆19Feb 8, 2011Updated 15 years ago
- Examples of use of pig scripting languages capabilities☆39Aug 1, 2016Updated 9 years ago
- Cascalog for the Impatient☆41Aug 16, 2013Updated 12 years ago
- collection of modules to build distributed and reliable concurrent systems in Python.☆207Sep 14, 2013Updated 12 years ago