Approximate cardinality estimation with HyperLogLog, as a Hive function
☆42Dec 17, 2012Updated 13 years ago
Alternatives and similar repositories for hive-udf
Users that are interested in hive-udf are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- GeoIP Functions for hive☆49Oct 13, 2020Updated 5 years ago
- Protobuf input format and Serde support☆18Mar 2, 2013Updated 13 years ago
- ☆34Jan 13, 2019Updated 7 years ago
- Example using Grafana with Druid☆11Mar 27, 2015Updated 11 years ago
- Framework that makes processing arbitrary binary data in Hadoop easier☆22Apr 8, 2013Updated 13 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Memcached on YARN☆19Jun 2, 2014Updated 12 years ago
- distributed read/write filesystem in go, bound to local mountpoint using go-fuse☆25Oct 5, 2015Updated 10 years ago
- Helpful user defined fuctions / table generating functions for Hive☆102May 2, 2016Updated 10 years ago
- scala driver for launching Amazon EMR jobs☆40Feb 10, 2016Updated 10 years ago
- Implementation of 'Recordinality' cardinality estimation sketch with distinct value sampling☆56Aug 20, 2013Updated 12 years ago
- Kubernetes deployment of PrestoDB, Hive Metastore, and Minio S3-standard object store☆17Oct 20, 2022Updated 3 years ago
- oozie designer and job management system☆22Sep 25, 2012Updated 13 years ago
- Bloomfilter support for Facebook Presto (prestodb.io)☆26Jul 7, 2022Updated 4 years ago
- Hadoop log aggregator and dashboard☆190Oct 29, 2013Updated 12 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Some Spark implementations of clustering algorithms.☆19Nov 13, 2018Updated 7 years ago
- Remedy small files by combining them into larger ones.☆196Jul 1, 2022Updated 4 years ago
- Stream summarizer and cardinality estimator.☆2,264Nov 28, 2019Updated 6 years ago
- Spark Tutorial at the University of Maryland☆37Oct 24, 2014Updated 11 years ago
- ☆17Jun 27, 2013Updated 13 years ago
- Applied Parallel Computing tutorial material for PyCon 2013 (Minesh Amin, Ian Ozsvald)☆17Apr 2, 2013Updated 13 years ago
- Distributed SQL base Realtime Streaming Computation Framework On Apache Storm, Spark☆12Mar 14, 2016Updated 10 years ago
- A NoSql database designed for maximum plugablitily and configurability.☆18Jul 30, 2025Updated last year
- Combination of Dockerized Hortonworks projects and other Hadoop ecosystem components☆10Oct 11, 2019Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- MapReduce examples☆20Nov 18, 2011Updated 14 years ago
- Code from my talk on Digital Signal Processing in Hadoop with Scalding☆15Oct 17, 2013Updated 12 years ago
- Collection of scripts for doing common transformations in machine learning☆21Dec 5, 2012Updated 13 years ago
- High-performance cryptography for the JVM☆18Jul 17, 2013Updated 13 years ago
- ☆16Sep 26, 2014Updated 11 years ago
- A Hadoop map reduce framework for Scala.☆15Apr 21, 2016Updated 10 years ago
- Unit test framework for hive and hive-service☆65Jun 29, 2022Updated 4 years ago
- Mahout Examples☆26Aug 2, 2016Updated 10 years ago
- Continuous Streaming SQL Queries for Flume☆96Dec 30, 2011Updated 14 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆16Apr 17, 2014Updated 12 years ago
- Ctrip Data Infrastructure team works for hue☆16Dec 10, 2014Updated 11 years ago
- HDFS Automatic Snapshot Service for Linux☆11Oct 17, 2016Updated 9 years ago
- Hive I/O Library☆67Oct 28, 2021Updated 4 years ago
- Spec file, init scripts, and other stuff to build a kafka RPM.☆35Oct 14, 2016Updated 9 years ago
- Packaging for redhat and fedora style RPM installations, including init.d scripts, default configurations, and a .spec file for building …☆24Jul 18, 2012Updated 14 years ago
- Data Management + Feed Processing Platform over Hadoop☆27May 8, 2013Updated 13 years ago