OLD - impyla now developed at `cloudera/impyla`
☆23Apr 16, 2014Updated 12 years ago
Alternatives and similar repositories for impyla-old
Users that are interested in impyla-old are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Framework that makes processing arbitrary binary data in Hadoop easier☆22Apr 8, 2013Updated 13 years ago
- Useful reusable pipeline components for Crunch jobs☆27Feb 10, 2015Updated 11 years ago
- Python wrapper for the hadoop WebHDFS Rest API☆32Apr 11, 2015Updated 11 years ago
- Jega is a concurrent networking and cooperative multitasking library for Python3.☆27Jul 17, 2013Updated 13 years ago
- cascading.jruby build and execution tool☆16Sep 23, 2015Updated 11 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A Ruby toolkit for cloud-friendly ETL☆37Jul 29, 2016Updated 10 years ago
- HBase adapters for Cascading☆47Aug 9, 2009Updated 17 years ago
- A JRuby DSL for Cascading☆16Jan 11, 2015Updated 11 years ago
- Mediates connections between an active-only FTP client (such as ftp.exe) and a server and allow it to transfer data passively.☆10Mar 29, 2018Updated 8 years ago
- DuckDB extension for MySQL☆15Mar 17, 2024Updated 2 years ago
- A project that implements statistical methods for identifying anomalous files☆22Dec 10, 2014Updated 11 years ago
- Cloudera Maven Archetypes☆18Sep 7, 2011Updated 15 years ago
- Cascading.Multitool is a sed and grep command line tool for Apache Hadoop.☆21May 1, 2012Updated 14 years ago
- Tools for working with parquet, impala, and hive☆135Jan 4, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A tool for running concurrent multi-configuration experiments.☆12Jul 10, 2017Updated 9 years ago
- JDBC adapter for Cascading☆23Jun 27, 2009Updated 17 years ago
- An R package to gather, munge, and convert event datasets into temporal event-networks.☆11Mar 28, 2018Updated 8 years ago
- Utilities to use Avro files from Hadoop Map/Reduce jobs and Streaming☆26Sep 10, 2013Updated 13 years ago
- ☆20Dec 4, 2024Updated last year
- Materialize plugin for dbt☆12Jan 25, 2021Updated 5 years ago
- Allows you to see where(datanodes) that contain a file in HDFS☆17Mar 16, 2013Updated 13 years ago
- A Django-Tastypie client☆17Nov 22, 2013Updated 12 years ago
- Mahout vector encoding for pig☆53Nov 20, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Provides a simple archetype to create MapReduce jobs with Maven.☆24Dec 3, 2010Updated 15 years ago
- local development sandbox containers☆21Updated this week
- ☆29Nov 17, 2014Updated 11 years ago
- HDFS client library for C☆45Mar 19, 2024Updated 2 years ago
- An http/json interface for Cassandra☆33Mar 23, 2010Updated 16 years ago
- Variable support VMOD☆25Apr 29, 2016Updated 10 years ago
- Python Client for WebHDFS REST API☆42May 8, 2015Updated 11 years ago
- ☆26Aug 14, 2013Updated 13 years ago
- Library for indexing VCF files for random access searches by rsID☆17Mar 2, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Exploration of spark streaming based on the BigData.be project 2☆15Sep 2, 2013Updated 13 years ago
- ☆14Apr 17, 2025Updated last year
- Python modules and scripts for working with Concrete, a data serialization format for NLP☆21Sep 1, 2026Updated last month
- Unix tee, but for Kinesis streams☆12Oct 19, 2021Updated 4 years ago
- The Hadoop GP Toolbox provides tools to exchange features between a Geodatabase and Hadoop and run Hadoop workflow jobs.☆75Jan 3, 2019Updated 7 years ago
- Slinky, a high-performance web crawler / text analytics in Python, Redis, Hadoop, R, Gephi☆40Aug 30, 2010Updated 16 years ago
- Spark In MapReduce (SIMR) - launching Spark applications on existing Hadoop MapReduce infrastructure☆44Mar 9, 2022Updated 4 years ago