Coding exercises for Apache Spark
☆103Jun 4, 2015Updated 11 years ago
Alternatives and similar repositories for spark-exercises
Users that are interested in spark-exercises are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code examples supporting the "Introduction to Apache Spark" video published by O'Reilly Media☆37Jul 1, 2022Updated 4 years ago
- Spark Tutorial at the University of Maryland☆37Oct 24, 2014Updated 11 years ago
- Sparking Using Java8☆17Feb 28, 2015Updated 11 years ago
- MapReduce examples☆20Nov 18, 2011Updated 14 years ago
- A WebGL distributed graph visualization library☆12Feb 26, 2015Updated 11 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Interactive Audience Analytics with Spark and HyperLogLog☆55Oct 14, 2015Updated 10 years ago
- Data and code for "Fast Data Applications with Spark and Python"☆25Sep 11, 2016Updated 9 years ago
- Application that visualizes your google location history in form of a heatmap using Spark to aggregate the data.☆12Feb 19, 2015Updated 11 years ago
- Ansible Role to install a Hadoop Cluster☆10Sep 21, 2020Updated 5 years ago
- Examples for Fast Data Processing with Spark☆59Sep 10, 2013Updated 12 years ago
- Elastic Search on Spark☆110Oct 21, 2014Updated 11 years ago
- ☆12Aug 29, 2015Updated 10 years ago
- Machine Learning for Cascading☆85Jun 12, 2015Updated 11 years ago
- Supporting content (slides and exercises) for the Addison-Wesley (Pearson) video series covering best practices for developing scalable S…☆68Jan 8, 2016Updated 10 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Slinky, a high-performance web crawler / text analytics in Python, Redis, Hadoop, R, Gephi☆40Aug 30, 2010Updated 15 years ago
- Application of Blockchain in Crop Farming and Crop Supply☆10May 15, 2018Updated 8 years ago
- ☆14Aug 24, 2021Updated 4 years ago
- HDFS compatible Distributed Filesystem backed Cassandra☆25Sep 17, 2015Updated 10 years ago
- ☆11Jul 30, 2014Updated 12 years ago
- BerkeleyX: CS100.1x, Introduction to Big Data with Apache Spark☆10Jul 27, 2015Updated 11 years ago
- [Deprecated] Docker image to run an out-of-the-box Memcached server☆11Mar 31, 2017Updated 9 years ago
- Solr on YARN prototype☆18Nov 14, 2014Updated 11 years ago
- Utilities for writing tests that use Apache Spark.☆24Dec 29, 2018Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Locality Sensitive Hashing for Apache Spark☆198Nov 1, 2016Updated 9 years ago
- LDA Analysis of the Twitter feed of @josephmisiti☆11Jul 1, 2014Updated 12 years ago
- tutorials and samples that show you how get the most out of IBM Analytics for Apache Spark☆78Mar 16, 2018Updated 8 years ago
- Tidy Data in Python Mini-Course by Vincent Lan [OPEN]☆10Jun 29, 2017Updated 9 years ago
- Tutorial on parsing Enron email to Avro and then explore the email set using Spark.☆52Mar 25, 2026Updated 4 months ago
- PySpark + Scikit-learn = Sparkit-learn☆1,151Dec 31, 2020Updated 5 years ago
- Practice coding by solving problems posted to careercup☆19Feb 28, 2012Updated 14 years ago
- db2ixf is a python package with a CLI that simplifies the parsing and processing of IBM Integration eXchange Format (IXF) files.☆16Jan 27, 2026Updated 6 months ago
- Application to securely map users on a multi tenant Amazon EMR cluster to different IAM Roles and then assume the mapped Role.☆24Oct 24, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Everyday Analytics and Visualization - JuliaCon 2015☆10Sep 25, 2015Updated 10 years ago
- Complete Pipeline Training at Big Data Scala By the Bay☆71Oct 27, 2015Updated 10 years ago
- Spark library for doing exploratory data analysis in a scalable way☆43Jan 17, 2016Updated 10 years ago
- ☆10Sep 16, 2016Updated 9 years ago
- Sparkling Pandas☆362Jul 6, 2023Updated 3 years ago
- Docker container capable of running an iPython notebook server, for "Just Enough Math"☆16Mar 31, 2023Updated 3 years ago
- SolrCloud Rebalance API Documentation☆13Jul 18, 2016Updated 10 years ago