A Python wrapper for Cascading
☆220Dec 30, 2019Updated 6 years ago
Alternatives and similar repositories for pycascading
Users that are interested in pycascading are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Elephant Twin LZO uses Elephant Twin to create LZO block indexes☆15Jun 13, 2012Updated 14 years ago
- simple simhashing in hadoop with cascading☆33May 9, 2011Updated 15 years ago
- hRaven collects run time data and statistics from MapReduce jobs in an easily queryable format☆129Jan 14, 2022Updated 4 years ago
- A grouping of Apache Pig examples.☆65Oct 13, 2020Updated 5 years ago
- A reporistory of User-defined functions for Apache Pig☆16Sep 20, 2010Updated 15 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Twitter's collection of LZO and Protocol Buffer-related Hadoop, Pig, Hive, and HBase code.☆1,133Apr 10, 2023Updated 3 years ago
- A Scala API for Cascading☆3,524May 28, 2023Updated 3 years ago
- Python module that allows one to easily write and run Hadoop programs.☆1,030Jan 9, 2018Updated 8 years ago
- All the Cascading taps you need and love.☆39Mar 4, 2019Updated 7 years ago
- The Scalding tutorial as a standalone SBT project☆51Oct 16, 2017Updated 8 years ago
- Common metadata layer for Hadoop's Map Reduce, Pig, and Hive☆77Feb 17, 2011Updated 15 years ago
- Elephant Twin is a framework for creating indexes in Hadoop☆98Oct 12, 2020Updated 5 years ago
- A platform for visualization and real-time monitoring of data workflows☆1,169Jan 22, 2020Updated 6 years ago
- Materials for Cascalog workshop☆18Sep 17, 2011Updated 14 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Streaming MapReduce with Scalding and Storm☆2,122Jan 19, 2022Updated 4 years ago
- A C interface to watchman☆47Jun 7, 2019Updated 7 years ago
- Framework that makes processing arbitrary binary data in Hadoop easier☆22Apr 8, 2013Updated 13 years ago
- Lightning-fast cluster computing in Java, Scala and Python.☆1,418Apr 8, 2014Updated 12 years ago
- Hadoop library for large-scale data processing, now an Apache Incubator project☆581Jul 8, 2014Updated 12 years ago
- Machine Learning for Cascading☆85Jun 12, 2015Updated 11 years ago
- Cascading is a feature rich API for defining and executing complex and fault tolerant data processing flows locally or on a cluster.☆355Apr 8, 2025Updated last year
- A Python client for the HBase Avro interface☆49Feb 1, 2016Updated 10 years ago
- Python MapReduce library written in Cython. Visit us in #hadoopy on freenode. See the link below for documentation and tutorials.☆243Jan 8, 2016Updated 10 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Optimized joins using bloom filters on Hadoop via Cascading.☆22Sep 25, 2009Updated 16 years ago
- Data processing on Hadoop without the hassle.☆1,373May 18, 2023Updated 3 years ago
- scalding powered machine learning☆109Nov 18, 2014Updated 11 years ago
- Watches files and records, or triggers actions, when they change.☆10Aug 26, 2015Updated 11 years ago
- cascading_ext is a collection of tools built on top of the Cascading platform which make it easy to build, debug, and run simple and high…☆58Feb 25, 2026Updated 6 months ago
- Bulk loading for elastic search☆186Dec 16, 2023Updated 2 years ago
- Main Repo☆15Jun 24, 2010Updated 16 years ago
- cascading schemes and taps for JDBC☆27Jun 15, 2016Updated 10 years ago
- Hadoop Input and Ouput formats for MongoDB☆29Nov 15, 2011Updated 14 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A simple key/value store for indexing meta data on JSON-type fields☆46Apr 25, 2011Updated 15 years ago
- A simple standalone Finagle stats viewer☆22Aug 28, 2013Updated 13 years ago
- Scala extensions for Storm☆132Jun 7, 2019Updated 7 years ago
- Utilities for working with Hadoop and Cascading☆19Feb 8, 2011Updated 15 years ago
- Examples of use of pig scripting languages capabilities☆39Aug 1, 2016Updated 10 years ago
- Cascalog for the Impatient☆41Aug 16, 2013Updated 13 years ago
- collection of modules to build distributed and reliable concurrent systems in Python.☆207Sep 14, 2013Updated 12 years ago