PySpark Cassandra brings back the fun in working with Cassandra data in PySpark.
☆79Jul 20, 2017Updated 9 years ago
Alternatives and similar repositories for pyspark-cassandra
Users that are interested in pyspark-cassandra are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- pyspark-cassandra is a Python port of the awesome @datastax Spark Cassandra connector. Compatible w/ Spark 2.0, 2.1, 2.2, 2.3 and 2.4☆68Jun 8, 2026Updated 2 months ago
- Python based data warehouse solution for the Lambda Architecture.☆14Jun 24, 2015Updated 11 years ago
- Utilities and examples to asssist in working with PySpark and Cassandra.☆35Feb 26, 2015Updated 11 years ago
- SigOpt's public R client☆13Aug 22, 2023Updated 3 years ago
- Cassandra in Docker☆127Jul 22, 2015Updated 11 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An R-like GLM package for Apache Spark☆10Aug 6, 2015Updated 11 years ago
- prototype editor for creating runnable network graphs☆12Mar 31, 2017Updated 9 years ago
- A fast, streaming-friendly, type-safe, pure-Scala MessagePack library. Supercharge your microservices today!☆60Jun 6, 2021Updated 5 years ago
- A lightweight library for reactivity☆20Oct 5, 2025Updated 10 months ago
- Unified interface for local and distributed ndarrays☆157Oct 13, 2018Updated 7 years ago
- A command line tool for Spark packages☆19Mar 30, 2023Updated 3 years ago
- ☆11Dec 26, 2022Updated 3 years ago
- Index and query k-mer matrices in BGZF☆12Apr 30, 2018Updated 8 years ago
- A place for my data sciencey related projects☆19Jul 16, 2019Updated 7 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆23Nov 21, 2018Updated 7 years ago
- Deploy Dask on Marathon☆10Feb 6, 2017Updated 9 years ago
- A demo explaining how to use Zeppelin notebook to access Apache Cassandra data via Apache Spark or CQL language☆17Mar 20, 2021Updated 5 years ago
- Code to allow running BIDMach on Spark including HDFS integration and lightweight sparse model updates (Kylix).☆16Jul 23, 2020Updated 6 years ago
- ☆12Sep 25, 2015Updated 10 years ago
- Open Automation Source Integrated Suite or (OASIS) is an open source plugin for FitNesse that enables the user to perform thick, thin, mo…☆15Apr 16, 2015Updated 11 years ago
- Apache Cassandra Health Check Tooling☆12Jun 11, 2026Updated 2 months ago
- Parallel Genomic Analysis Toolkit☆14Feb 11, 2019Updated 7 years ago
- REST job server for Apache Spark☆2,836Mar 3, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- python client library☆10Feb 15, 2017Updated 9 years ago
- functionstest☆33Oct 25, 2016Updated 9 years ago
- A mesos plugin for Relay that lets you auto-scale the number of currently running instances of a bash command☆38May 16, 2021Updated 5 years ago
- A Unix backend for Ibis☆11Jul 11, 2024Updated 2 years ago
- ☆24Dec 10, 2019Updated 6 years ago
- Tool for tagging FLV files☆12Nov 2, 2015Updated 10 years ago
- Walk the working directory towards the root recursively☆17Jul 31, 2025Updated last year
- Open source analytics platform powered by Apache Cassandra, Spark, and Kafka☆34May 8, 2015Updated 11 years ago
- Adjacency-constrained hierarchical clustering of a similarity matrix☆16Nov 21, 2025Updated 9 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- PySpark + Scikit-learn = Sparkit-learn☆1,150Dec 31, 2020Updated 5 years ago
- A collection and conversion of WARN notices from California☆12May 13, 2016Updated 10 years ago
- Materials for PyData at Strata/Hadoop World San Jose 2015☆12Mar 9, 2015Updated 11 years ago
- Restricted Boltzmann Machines in R☆36Oct 1, 2020Updated 5 years ago
- Local-first federated analytics query engine using DuckDB.☆18Jan 30, 2025Updated last year
- Python interface to the stats.nba.com HTTP API.☆11Dec 26, 2022Updated 3 years ago
- Ansible playbooks to construct distributed computing environments☆62Jun 6, 2021Updated 5 years ago