Distributed SQL Query Engine in Python using Ray
☆245Oct 2, 2024Updated last year
Alternatives and similar repositories for ray-sql
Users that are interested in ray-sql are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Experimental DataFusion Optimizer☆52Jun 9, 2023Updated 3 years ago
- Unofficial rust implementation of Apache Iceberg with integration for Datafusion☆241Updated this week
- Batteries included CLI, TUI, and server implementations for DataFusion.☆199Jul 13, 2026Updated last week
- Apache DataFusion Ray☆230May 15, 2026Updated 2 months ago
- Apache DataFusion Ballista Distributed Query Engine☆2,093Updated this week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Apache DataFusion Comet Spark Accelerator☆1,230Updated this week
- ☆34May 9, 2025Updated last year
- SQL Benchmark derived from TPC-H☆11May 20, 2023Updated 3 years ago
- Apache Iceberg☆1,348Updated this week
- A cross platform way to express data transformation, relational algebra, standardized record expression and plans.☆1,535Updated this week
- Query Plan Markup Language☆45Jan 18, 2024Updated 2 years ago
- Apache DataFusion Python Bindings☆594Updated this week
- Pure Rust Iceberg Implementation☆162Aug 13, 2024Updated last year
- SQLBench Runners☆13Dec 17, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The Auron accelerator for distributed computing framework (e.g., Spark) leverages native vectorized execution to accelerate query process…☆1,778Updated this week
- Apache DataFusion SQL Query Engine☆9,000Updated this week
- GlareDB: A light and fast SQL database for analytics☆1,017Nov 14, 2025Updated 8 months ago
- RayDP provides simple APIs for running Spark on Ray and integrating Spark with AI libraries.☆374Jun 10, 2026Updated last month
- Fastest and safest Rust implementation of parquet. `unsafe` free. Integration-tested against pyarrow☆385Jul 31, 2024Updated last year
- CMU-DB's Cascades optimizer framework☆405Jan 6, 2025Updated last year
- Distributed SQL Engine in Python using Dask☆411Aug 29, 2024Updated last year
- Making data lake work for time series☆1,192Aug 21, 2024Updated last year
- Rust lib to read from Apache ORC☆18Jun 9, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A collection of RBIR projects and posts for anyone interested in joining this journey.☆327Updated this week
- Exoshuffle-CloudSort☆30Mar 2, 2023Updated 3 years ago
- A leightweight UI for Lakekeeper☆19Updated this week
- Cache server :)☆32Sep 5, 2023Updated 2 years ago
- Analytical database for data-driven Web applications 🪶☆514Feb 25, 2025Updated last year
- A composable and fully extensible C++ execution engine library for data management systems.☆4,173Updated this week
- Boring Data Tool☆244Mar 21, 2024Updated 2 years ago
- Quickly view your data☆363Jul 9, 2026Updated last week
- Tools for generating TPC-* datasets☆33Jun 23, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A native storage format for apache arrow☆83Oct 18, 2023Updated 2 years ago
- ☆15Feb 4, 2026Updated 5 months ago
- A reader that buffers ranged calls☆12May 17, 2022Updated 4 years ago
- Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random access, vector index, and data ve…☆6,825Updated this week
- A native Rust library for Delta Lake, with bindings into Python☆3,265Updated this week
- Gluten is a middle layer responsible for offloading JVM-based SQL engines' execution to native engines.☆1,576Updated this week
- New and extensible file format for storage of large columnar datasets.☆728Updated this week