A tool to benchmark L (loading) workloads within ETL workloads
☆35Aug 24, 2026Updated last month
Alternatives and similar repositories for lake-loader
Users that are interested in lake-loader are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Quanton Operator is a Kubernetes operator that extends kubeflow/spark-operator to run Apache Spark jobs using the Quanton compute engine …☆23Updated this week
- Monitoring and insights on your data lakehouse tables☆34Sep 17, 2026Updated last week
- Hudi Demo Notebook☆11Mar 5, 2024Updated 2 years ago
- 2025 Vibe Coding Session☆15Sep 29, 2025Updated 11 months ago
- Apache XTable (incubating) is a cross-table converter for lakehouse table formats that facilitates interoperability across data processin…☆1,250Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An example of a multiple workspace deployment with reusable modules.☆13May 28, 2025Updated last year
- A repo of Java examples using Apache Flink with flink-connector-kafka☆10Mar 10, 2026Updated 6 months ago
- DevOps Projects undertaken in Roadmap.sh☆17Aug 19, 2025Updated last year
- 思科vpn客户端☆13Nov 24, 2016Updated 9 years ago
- ☆13Mar 29, 2023Updated 3 years ago
- Apache Spark Connect Client for Rust (Rust core + Python wrapper)☆45Sep 11, 2026Updated 2 weeks ago
- Python client code for Apache HBase 1.x using Kerberos authentication over SASL☆11Jan 9, 2017Updated 9 years ago
- ☆30Dec 4, 2024Updated last year
- ☆16Jun 22, 2015Updated 11 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Expect script to change login passwords using SSH☆17Jan 14, 2024Updated 2 years ago
- ☆24Dec 17, 2025Updated 9 months ago
- Contains tools for analyzing time-series data.☆11May 8, 2013Updated 13 years ago
- Instructions for setting up Kerberos, Zookeeper, and Kafka with SASL☆16Jan 22, 2018Updated 8 years ago
- Upserts, Deletes And Incremental Processing on Big Data.☆6,274Updated this week
- A simple, yet customizable floating clock for Linux. A lightweight alternative to @kostoskistefan/hoverclock-qt☆17Aug 11, 2025Updated last year
- Fuzzing compression libraries☆20Jan 10, 2016Updated 10 years ago
- Monitoring Databricks using Prometheus, Grafana and Pyroscope☆28Jul 29, 2025Updated last year
- Prometheus Metrics Reporter for Apache Kafka server and client components☆22Sep 15, 2026Updated last week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Collection of Terraform code for AWS labs☆14Mar 5, 2026Updated 6 months ago
- Integration of Iceberg table management into Spark SQL☆11Jan 21, 2020Updated 6 years ago
- Simulating monty hall problem with python☆11May 22, 2022Updated 4 years ago
- An open-source storage framework that enables building a Lakehouse architecture with compute engines including Spark, PrestoDB, Flink, Tr…☆10Feb 10, 2023Updated 3 years ago
- Super Mario is a legendary game we all cherish! In this project, we will deploy Super Mario on Amazon EKS (Elastic Kubernetes Service) us…☆14Feb 3, 2026Updated 7 months ago
- A Tez dev-setup for HDP2 sandbox☆21Mar 2, 2023Updated 3 years ago
- Drop-in replacement for Apache Spark UI☆491Aug 20, 2026Updated last month
- ☆18Updated this week
- A tiny servlet container using netty.☆14Feb 17, 2014Updated 12 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Oct 20, 2021Updated 4 years ago
- A methodology to allow customers to ingest, store, and analyze gameplay telemetry in near real-time.☆13Feb 28, 2020Updated 6 years ago
- ☆15Feb 1, 2016Updated 10 years ago
- Repository to go along with the paper "Plumber: Diagnosing and Removing Performance Bottlenecks in Machine Learning Data Pipelines"☆10Mar 31, 2022Updated 4 years ago
- ☆13Jul 14, 2026Updated 2 months ago
- Build, Test and Deploy ETL solutions using AWS Glue and AWS CDK based CI/CD pipelines☆44Oct 20, 2022Updated 3 years ago
- This checklist aims to be an exhaustive list of all elements you should consider when using Amazon Redshift.☆15Sep 21, 2020Updated 6 years ago