A library for loadling Thrift data into Spark SQL
☆43Mar 4, 2023Updated 3 years ago
Alternatives and similar repositories for airbnb-spark-thrift
Users that are interested in airbnb-spark-thrift are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A sink to save Spark Structured Streaming DataFrame into Hive table☆23May 7, 2018Updated 8 years ago
- This is a simple CEP Engine leveraging the Kafka Streams platform☆16Apr 25, 2017Updated 9 years ago
- Utilities for writing tests that use Apache Spark.☆24Dec 29, 2018Updated 7 years ago
- Scala Goodies☆41Jun 9, 2016Updated 10 years ago
- HBase Indexer - indexing HBase to Solr 5.x and higher☆13Oct 27, 2017Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A simple project that trains an OpenNLP Named Entity Recognition model to identify ingredients in a recipe.☆14Oct 30, 2016Updated 9 years ago
- An Apache Mesos Framework that allows for replaying load over and over and over (and over) again☆10Aug 10, 2015Updated 11 years ago
- Spark Modularized View☆43May 25, 2020Updated 6 years ago
- Apache flink☆21May 15, 2026Updated 3 months ago
- Dynamically loads bundled JNI libraries based on the runtime platform.☆10Dec 19, 2014Updated 11 years ago
- Scripts for running Apache Kafka on Mesosphere's Marathon☆14Dec 6, 2015Updated 10 years ago
- A Spark metrics sink that pushes to InfluxDb☆51Jan 14, 2021Updated 5 years ago
- phData Pulse application log aggregation and monitoring☆13Apr 13, 2020Updated 6 years ago
- Spark CEP is an extension of Spark Streaming to support SQL-based query processing☆58Apr 12, 2017Updated 9 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- 텐서플로우 뽀개기 - python 코드를 R로 변경하기☆14Sep 5, 2017Updated 9 years ago
- Create and use de-identified research databases. Preprocess, extract text, anonymise/de-identify, link, apply natural language processing…☆24Aug 3, 2026Updated last month
- 用GDB调试 Go语言☆10Apr 13, 2015Updated 11 years ago
- Simply add a vagrant based jupyter, anaconda environment to your python project☆13Aug 9, 2017Updated 9 years ago
- 各种有用的web api 基于Golang, Python(tornado django scrapy gevent)☆10Feb 19, 2016Updated 10 years ago
- Experiments with the GDELT dataset and Cassandra schemas.☆25Feb 9, 2016Updated 10 years ago
- ☆14Aug 19, 2023Updated 3 years ago
- Spark MLlib code optimized to efficiently support sparse data☆51Dec 22, 2016Updated 9 years ago
- Kafka Streams DSL vs Processor API☆16Nov 26, 2017Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A set of tools to ease working with Zookeeper and Kafka.☆23Jan 22, 2016Updated 10 years ago
- spark structured streaming via HTTP communication☆18Jul 7, 2022Updated 4 years ago
- Sample code to help with Elastic Block Store automation with Elastic Volumes feature☆12Feb 24, 2017Updated 9 years ago
- A Real-Time Analytical Processing (RTAP) example using Spark/Shark☆51Feb 21, 2014Updated 12 years ago
- game server based on wentby frame☆11Jun 12, 2020Updated 6 years ago
- Apache-Spark based Data Flow(ETL) Framework which supports multiple read, write destinations of different types and also support multiple…☆26Jun 7, 2021Updated 5 years ago
- Embedded Kafka for testing and quick prototyping.☆14Apr 19, 2016Updated 10 years ago
- Library for organizing batch processing pipelines in Apache Spark☆43Jan 4, 2017Updated 9 years ago
- A small project to report offset lag for Kafka Consumer Groups via Burrow.☆32Nov 5, 2019Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Kubernetes operator for Trino. Deploy and run Trino coordinators and workers with the Stackable Data Platform (SDP).☆63Updated this week
- ☆13Dec 23, 2020Updated 5 years ago
- Spark package to "plug" holes in data using SQL based rules ⚡️ 🔌☆28May 15, 2020Updated 6 years ago
- Apache flink☆24May 15, 2026Updated 3 months ago
- Spark Structured Streaming State Tools☆35Jul 3, 2020Updated 6 years ago
- Docker based Netflix proxy☆18Sep 6, 2014Updated 12 years ago
- 🔬OpenTracing instrumentation for the Sanic Web Framework☆12Aug 6, 2022Updated 4 years ago