☆245Oct 7, 2019Updated 6 years ago
Alternatives and similar repositories for Apache-Kafka-poc-and-notes
Users that are interested in Apache-Kafka-poc-and-notes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆130Apr 8, 2017Updated 9 years ago
- Educational notes,Hands on problems w/ solutions for hadoop ecosystem☆87Jan 22, 2019Updated 7 years ago
- Flume-to-Spark-Streaming Log Parser☆23Jun 3, 2016Updated 10 years ago
- Examples of Spark 2.0☆213Aug 11, 2021Updated 5 years ago
- The Internals of Apache Spark☆1,551Jul 18, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Spark Gotchas. A subjective compilation of the Apache Spark tips and tricks☆358Jun 6, 2017Updated 9 years ago
- ☆313Nov 26, 2018Updated 7 years ago
- Qubole Streaminglens tool for tuning Spark Structured Streaming Pipelines☆17Jan 21, 2020Updated 6 years ago
- My MSc on Data Science final project. This is a library for Data Pre-processing Algorithms for Streaming in Flink (DPASF)☆18Jul 1, 2019Updated 7 years ago
- Sample processing code using Spark 2.1+ and Scala☆51Jun 28, 2020Updated 6 years ago
- Structured Streaming Machine Learning example with Spark 2.0☆96Apr 24, 2017Updated 9 years ago
- Example projects for using Spark and Cassandra With DSE Analytics☆59Oct 10, 2025Updated 11 months ago
- This project enables you to use spring inside of a spark application.☆11May 6, 2015Updated 11 years ago
- Contain Interview Questions Solutions☆12May 18, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ## Auto-archived due to inactivity. ## Simple JVM Profiler Using StatsD and Other Metrics Backends☆15Oct 3, 2023Updated 2 years ago
- Nested Data (JSON/AVRO/XML) Parsing and Flattening in Spark☆16Jan 22, 2024Updated 2 years ago
- Apache Spark (Scala, PySpark, SparkR) Code, Tricks, and References☆69Jan 21, 2019Updated 7 years ago
- A library you can include in your Spark job to validate the counters and perform operations on success. Goal is scala/java/python support…☆112Feb 1, 2018Updated 8 years ago
- Essential Spark extensions and helper methods ✨😲☆767Jun 22, 2026Updated 3 months ago
- A boilerplate for writing PySpark Jobs☆393Jan 21, 2024Updated 2 years ago
- ACID Data Source for Apache Spark based on Hive ACID☆97Jul 7, 2021Updated 5 years ago
- This repository houses the Query It! experience.☆11Apr 29, 2020Updated 6 years ago
- Some AWS EMR examples☆16Jan 18, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Build end-to-end Machine Learning pipeline to predict accessibility of playgrounds in NYC☆15Jul 9, 2020Updated 6 years ago
- Repository used for Spark Trainings☆54Apr 21, 2023Updated 3 years ago
- Self-contained examples of Apache Spark streaming integrated with Apache Kafka.☆196Apr 15, 2018Updated 8 years ago
- Examples for High Performance Spark☆537May 3, 2026Updated 4 months ago
- Udacity Data Engineer Nanodegree - Capstone project☆11Dec 19, 2019Updated 6 years ago
- Base classes to use when writing tests with Spark☆1,555Aug 31, 2026Updated 3 weeks ago
- ☆38May 27, 2025Updated last year
- Apache Spark™ and Scala Workshops☆264Jul 29, 2024Updated 2 years ago
- The documentation for the Clustergrammer project☆10Oct 21, 2020Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- This tutorial provides a quick introduction to using Spark☆58Mar 31, 2016Updated 10 years ago
- A library for reading data from Amzon S3 with optimised listing using Amazon SQS using Spark SQL Streaming ( or Structured streaming).☆19Apr 20, 2024Updated 2 years ago
- Examples of Spark 3.0☆44Nov 11, 2020Updated 5 years ago
- Developing Spark External Data Sources using the V2 API☆49Apr 29, 2018Updated 8 years ago
- CDM conversion of MIMIC dataset.☆17Jun 19, 2016Updated 10 years ago
- A small Python module to parse RFC5424-formatted Syslog messages☆37Oct 17, 2025Updated 11 months ago
- A tutorial on Apache Spark Unit Testing☆38Jan 27, 2016Updated 10 years ago