IndexTables is an open-table format for Apache Spark that enables fast retrieval and full-text search across large-scale data. It integrates seamlessly with Spark SQL, allowing you to combine powerful search capabilities with joins, aggregations, and standard SQL operations.
☆43Aug 3, 2026Updated last week
Alternatives and similar repositories for indextables_spark
Users that are interested in indextables_spark are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SIEM-to-Spark Transpiler☆44Mar 27, 2026Updated 4 months ago
- A reader that buffers ranged calls☆12May 17, 2022Updated 4 years ago
- Databricks Add-on for Splunk☆29Jan 15, 2026Updated 6 months ago
- The home of Floecat: A catalog of catalogs for open table formats☆89Updated this week
- Apache Spark Sandbox with patterns often focusing on Microsoft Fabric☆20Jun 21, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Java bindings for Apache DataFusion☆29Updated this week
- This project is a core-library for generate POJO's from asyncApi yaml contract.☆15Jul 1, 2026Updated last month
- A lightweight, production-ready Spring Boot library to standardize REST API responses and Global Exception Handling (RFC Standard). Inclu…☆17Feb 16, 2026Updated 5 months ago
- Unleash the performance potential of your Parquet files.☆55Feb 24, 2026Updated 5 months ago
- Integration of opentelemetry with the tracing crate☆27Aug 1, 2026Updated last week
- ☆16Jul 25, 2025Updated last year
- Log search engine on object storages☆21Jul 27, 2024Updated 2 years ago
- JavaWeb开发脚手架☆20Jul 28, 2026Updated last week
- Kafka Consumer Offset Monitoring☆21Jan 26, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Local code search for AI coding agents: a CLI and MCP server with hybrid keyword + semantic search and SQL relevance-ranked aggregation o…☆19Updated this week
- Open, Multi-modal Catalog for Data & AI, written in Rust☆86Sep 30, 2024Updated last year
- Spark* shuffle plugin for support shuffling data through a remote Hadoop-compatible file system, as opposed to vanilla Spark's local-dis…☆21Mar 15, 2024Updated 2 years ago
- Scala KairosDB driver☆15Updated this week
- Business Rule Engine☆22Jul 24, 2026Updated 2 weeks ago
- ☆14Jun 10, 2024Updated 2 years ago
- Notion to Postgres sync pipeline with Dagster orchestration☆19Jul 21, 2025Updated last year
- Using WASM to write UDFs in Apache Spark☆12Jun 3, 2024Updated 2 years ago
- An open-source, community-driven REST catalog for Apache Iceberg!☆30Jun 26, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A replicated state machine built atop S3.☆30May 12, 2026Updated 2 months ago
- Apache Hive Metastore in Standalone Mode With Docker☆14Jul 22, 2024Updated 2 years ago
- ☆30Dec 4, 2024Updated last year
- A console-based interface for Github code reviews☆18Jul 7, 2026Updated last month
- Minimal, exact vector search with metadata filtering. Think "Polars for vector search."☆34Jan 31, 2026Updated 6 months ago
- Spring Swing is a framework designed for building Spring-powered swing applications.☆15Jul 11, 2026Updated 3 weeks ago
- Advanced fold methods for Kotlin☆13Aug 1, 2026Updated last week
- A simple Rust crate to cache data both in-memory and on disk☆11Dec 26, 2021Updated 4 years ago
- Unusual CSV library for Java☆16Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A Gentle introduction to Machine Learning with Apache Spark☆11Mar 2, 2026Updated 5 months ago
- Lava — 让 Java 基础设施开发更简单、更安全、更一致☆15Apr 21, 2026Updated 3 months ago
- Client libraries to interface with Fluxzero Runtime☆18Updated this week
- JUnit5 extension and helpers for writing tests parameterised over Kafka clusters☆16Updated this week
- Lightweight persistence layer framework, support interface mapping, dynamic sql and sql file manager.☆17Updated this week
- Visits sessionization pipeline used for the talk☆13May 28, 2024Updated 2 years ago
- Apache Web Services - XmlSchema☆17Updated this week