HadoopOffice - Analyze Office documents using the Hadoop ecosystem (Spark/Flink/Hive)
☆63Sep 29, 2025Updated 10 months ago
Alternatives and similar repositories for hadoopoffice
Users that are interested in hadoopoffice are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Spark datasource for the HadoopOffice library☆36Sep 29, 2025Updated 10 months ago
- A Spark plugin for reading and writing Excel files☆523Updated this week
- 【易车】- Spark、flink、HBase、Hive、flume集成了一些Hadoop的原生api的一些demo(如HDFS、MapReduce:目前就这两个);同时测试一些异常功能☆16Apr 4, 2019Updated 7 years ago
- Type-safe SQL builder for Scala☆31Jun 4, 2026Updated 2 months ago
- painless Github client☆14Sep 4, 2015Updated 10 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- A Spark data source for reading Microsoft Excel files☆13Jul 1, 2024Updated 2 years ago
- Automated TPC-DS and TPC-H benchmark for Apache Hive LLAP☆10Jul 18, 2022Updated 4 years ago
- low-level helpers for Apache Spark libraries and tests☆16Dec 29, 2018Updated 7 years ago
- Cloudera Manager parcel and CSD to manage Cassandra NoSQL database☆14Nov 16, 2016Updated 9 years ago
- A JBoss BPM Suite with JBoss Fuse demo project around a travel agency booking workflow and migration to micro-services.☆15Mar 23, 2016Updated 10 years ago
- Simple retry functionality for Scala☆21Dec 11, 2021Updated 4 years ago
- Scala API for Apache Spark SQL high-order functions☆15Aug 4, 2023Updated 3 years ago
- Java Kerberos Utilities☆13Feb 7, 2016Updated 10 years ago
- Spark on Kudu up and running samples☆10Jan 29, 2017Updated 9 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Demo project for bluetooth lamp☆13Feb 13, 2026Updated 5 months ago
- Distributed Web Crawler, Parser and Search Engine.☆10Jun 16, 2016Updated 10 years ago
- Detect flaky tests with sbt☆38Aug 16, 2023Updated 2 years ago
- Apache Shiro plugin for Stormpath☆34May 1, 2017Updated 9 years ago
- ACID Data Source for Apache Spark based on Hive ACID☆97Jul 7, 2021Updated 5 years ago
- A set of widgets for Python's Orange Machine Learning to work with Apache Spark ML☆15Dec 24, 2016Updated 9 years ago
- MySQL Binlog based Changed Data Capture☆11Apr 26, 2017Updated 9 years ago
- Swimlane graphs for Hive, SparkSQL, and Presto based on Ganglia resource graphs☆13Feb 13, 2017Updated 9 years ago
- Repository for my article: Google OAuth 2.0 – Testing with Curl☆16Jun 18, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Cloud Spanner Connector for Apache Spark☆18Updated this week
- Utilities for writing tests that use Apache Spark.☆24Dec 29, 2018Updated 7 years ago
- React Native module for lightweight universal authentication using Keycloak☆23Feb 8, 2023Updated 3 years ago
- Schema Registry integration for Apache Spark☆40Nov 16, 2022Updated 3 years ago
- ☆15Jun 19, 2016Updated 10 years ago
- Spark Structured Streaming State Tools☆34Jul 3, 2020Updated 6 years ago
- Example for a Gradle project for LEGO Mindstorm EV3 and Lejos☆13Nov 28, 2021Updated 4 years ago
- Google Maps geocoding library for Scala☆12Oct 12, 2019Updated 6 years ago
- Scrapy exporter for Big Data formats☆16Mar 10, 2026Updated 5 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A MongoDB connector for Apache Flink prior to 1.16. For Flink 1.16+, please use the official MongoDB connector.☆49Jul 27, 2023Updated 3 years ago
- Deriving Spark DataFrame schemas from case classes☆44Jun 24, 2024Updated 2 years ago
- Coursera's big data course with Scala and Spark☆16Mar 29, 2017Updated 9 years ago
- Generate literate-style markdown docs from your sources☆59Jan 30, 2018Updated 8 years ago
- 这是一个扩展的flink-connector-jdbc,相比于官方,该版本新增了对clickhouse、phoenix的支持,后面将会继续改造以支持更多的jdbc连接☆16May 13, 2022Updated 4 years ago
- Shaded version of Apache Hive for Trino☆11Aug 13, 2024Updated last year
- kamon netty integration☆10Aug 30, 2020Updated 5 years ago