Utilities and examples to asssist in working with PySpark and Cassandra.
☆35Feb 26, 2015Updated 11 years ago
Alternatives and similar repositories for pyspark-cassandra
Users that are interested in pyspark-cassandra are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆34May 24, 2014Updated 12 years ago
- Packer Template to build a AWS Apache Zookeeper AMI☆14Jan 3, 2022Updated 4 years ago
- Python port of CoalescingRingBuffer from LMAXCollections☆16Aug 11, 2013Updated 13 years ago
- Offline Elasticsearch index generator☆26Apr 29, 2021Updated 5 years ago
- ☆17Jun 14, 2014Updated 12 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A collection of ipython notebooks I've made for various projects☆11Mar 23, 2015Updated 11 years ago
- Tutorial for performing queries on the Wikipedia API for social network analysis☆48Oct 9, 2014Updated 11 years ago
- Gobbling it's way through your pcap files☆24Oct 23, 2014Updated 11 years ago
- PySpark Cassandra brings back the fun in working with Cassandra data in PySpark.☆79Jul 20, 2017Updated 9 years ago
- Distributed and live analysis of log data using python, supervisord and redis. Suitable for repositories or sites that wish to count stat…☆23Apr 12, 2010Updated 16 years ago
- A visual stat tool for CrateDB clusters☆12Jun 1, 2018Updated 8 years ago
- Parallelize network dependent tasks in python.☆27Dec 14, 2014Updated 11 years ago
- A Django app to make simple, opinionated reports☆18Apr 21, 2014Updated 12 years ago
- zbx-statsd is a clone of Etsy's statsd and Steve Ivy's py-statsd designed to work with Zabbix (http://zabbix.com).☆19Jun 15, 2012Updated 14 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Gobblin is a distributed big data integration framework (ingestion, replication, compliance, retention) for batch and streaming systems.…☆11Jul 29, 2017Updated 9 years ago
- Terraform Module to create a Apache Zookeeper cluster on AWS☆13Jan 3, 2022Updated 4 years ago
- Dummy variable generation with fit/transform capabilities☆23Aug 7, 2018Updated 8 years ago
- Grafana Prometheus exporter☆10Oct 17, 2017Updated 8 years ago
- Remember a 3x3 m² location anywhere in the world with just four words.☆22Jul 4, 2017Updated 9 years ago
- Replication materials for Bayesian measurement error model of dichotomous measures of democracy.☆16May 12, 2015Updated 11 years ago
- ☆16Jan 21, 2015Updated 11 years ago
- Real time and offline time series analysis with Spark, Spark Streaming and Storm☆21Oct 20, 2020Updated 5 years ago
- Packer Template to build a AWS Apache Cassandra AMI☆10Jan 3, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Sample Docker Compose files for running Apache Ambari☆11Oct 29, 2018Updated 7 years ago
- Materials for my talk at PyData Chicago 2016☆20May 25, 2017Updated 9 years ago
- Malware.lu configuration extractor☆26Mar 27, 2014Updated 12 years ago
- Adds a framework to enable Natural Language interactions in your Hubot scripts☆11Dec 6, 2016Updated 9 years ago
- This package implements a Hadoop metrics.spi.MetricContext for sending metric data to collectd☆15Aug 19, 2011Updated 15 years ago
- Implementation of an algorithm computing the nearest "N" neighbours to a vector, using a collection of hyperplane hashers.☆30Jul 17, 2015Updated 11 years ago
- Unix tee, but for Kinesis streams☆12Oct 19, 2021Updated 4 years ago
- lxml-cffi is a PyPy-friendly port of lxml, based on cffi☆21Oct 13, 2015Updated 10 years ago
- Jiraya - Simple Jira CLI☆17Dec 13, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A scala dsl for dataflow☆11Dec 31, 2014Updated 11 years ago
- Custom Alerts for Ambari server☆12Jul 27, 2015Updated 11 years ago
- A bridge to Apache Atlas for provenance metadata created in course of using Apache NiFi☆16Jan 2, 2023Updated 3 years ago
- Talk for PyMNtos on code generation in SymPy☆18Sep 18, 2014Updated 12 years ago
- Partitioned storage system based on blosc. **No longer actively maintained.**☆157Nov 21, 2016Updated 9 years ago
- Sukija is a program for indexing documents written in Finnish.☆19Apr 12, 2026Updated 5 months ago
- Tool to migrate Prometheus 1.x data directories to the 2.0 format.☆14Jan 18, 2018Updated 8 years ago