TPC-DS benchmark kit with some modifications/fixes
☆107Aug 13, 2024Updated 2 years ago
Alternatives and similar repositories for tpcds-kit
Users that are interested in tpcds-kit are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Use the TPC-DS benchmark to test Spark SQL performance☆185Apr 27, 2020Updated 6 years ago
- TPC-DS benchmark kit with some modifications/fixes☆365Apr 16, 2024Updated 2 years ago
- Parquet file generator☆22Apr 17, 2018Updated 8 years ago
- TPC-H queries in Apache Spark SQL using native DataFrames API☆99Jan 24, 2024Updated 2 years ago
- Patched version of dbgen☆34Feb 25, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Flink, Presto, Trino TPC-DS benchmark☆16Feb 20, 2023Updated 3 years ago
- TPC-DS benchmark kit with some modifications/additions☆10Nov 12, 2015Updated 10 years ago
- ☆23May 12, 2018Updated 8 years ago
- Gluten is a middle layer responsible for offloading JVM-based SQL engines' execution to native engines.☆1,600Updated this week
- Run TPC-DS against different databases including Hive, Spark SQL and IBM BigSQL☆14Jan 4, 2022Updated 4 years ago
- ☆391Jan 25, 2024Updated 2 years ago
- This is archive of SparkRDMA project. The new repository with RDMA shuffle acceleration for Apache Spark is here: https://github.com/Nvid…☆258May 13, 2019Updated 7 years ago
- Layered Multiple Functional Aggregate Optimization☆17Oct 8, 2020Updated 5 years ago
- Midas is a memory management system that efficiently and safely harvests idle memory for applications' soft state.☆12Oct 30, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Cloud Shuffle Service(CSS) is a general purpose remote shuffle solution for compute engines, including Spark/Flink/MapReduce.☆260May 12, 2024Updated 2 years ago
- Apache Celeborn is an elastic and high-performance service for shuffle and spilled data.☆1,066Updated this week
- collection of read materials☆18May 18, 2020Updated 6 years ago
- ☆42Jul 4, 2022Updated 4 years ago
- TPC-DS benchmarks☆73Jan 19, 2024Updated 2 years ago
- A serverless datalake project and framework based on AWS S3,Glue,Athena,MWAA and QuickSight. With a series of best practices, it guides y…☆16Nov 22, 2022Updated 3 years ago
- Uniffle is a high performance, general purpose Remote Shuffle Service.☆454Sep 10, 2026Updated last week
- Utility for benchmarking changes in Spark using TPC-DS workloads☆16Jun 3, 2021Updated 5 years ago
- This is a demo for pretty print ClickHouse variables☆19Mar 2, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Reproducing TPC-DS qualification/reference results☆37Aug 16, 2023Updated 3 years ago
- ☆10Jun 8, 2016Updated 10 years ago
- HiBench is a big data benchmark suite.☆1,485Dec 15, 2025Updated 9 months ago
- A repo for all spark examples using Rapids Accelerator including ETL, ML/DL, etc.☆171Updated this week
- A composable and fully extensible C++ execution engine library for data management systems.☆4,211Updated this week
- Implementing MISON by Microsoft in C++ as a test☆21Mar 1, 2018Updated 8 years ago
- TPC-H benchmark kit with some modifications/additions☆218Jul 20, 2022Updated 4 years ago
- SQL Benchmark derived from TPC-H☆11May 20, 2023Updated 3 years ago
- ☆26Apr 26, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Self-contained demo using Flink SQL and Debezium to build a CDC-based analytics pipeline. All you need is Docker!☆26May 11, 2021Updated 5 years ago
- A high-performance, scalable and efficient ShuffleManager plugin for Apache Spark, utilizing UCX communication layer☆51Oct 30, 2023Updated 2 years ago
- Demo of DuckDB Spark API implements. Same Pyspark code, but DuckDB under the hood☆15Nov 16, 2023Updated 2 years ago
- NVIDIA cuDF for Apache Spark plugin - accelerate Apache Spark with GPUs☆1,005Updated this week
- TPC-DS Kit for Impala☆170May 20, 2024Updated 2 years ago
- Scalable Structural Index Constructor for JSON Analytics☆27Oct 10, 2024Updated last year
- Star Schema Benchmark dbgen☆124Mar 11, 2024Updated 2 years ago