Yggdrasil: Faster Decision Trees Using Column Partitioning in Spark
☆30May 17, 2018Updated 8 years ago
Alternatives and similar repositories for yggdrasil
Users that are interested in yggdrasil are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Parquet Command-line Tools☆19Oct 26, 2016Updated 9 years ago
- Scripts for building Cloudera Manager parcel and CSD for Livy Spark Server☆21Oct 18, 2017Updated 8 years ago
- Distributed implementation of Robust PLSA using Spark☆12Apr 29, 2021Updated 5 years ago
- Gust is a set of GPU extensions for Breeze.☆32Apr 10, 2015Updated 11 years ago
- ☆20Jul 7, 2017Updated 9 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Spark Extension : ML transformers, SQL aggregations, etc that are missing in Apache Spark☆144Jan 26, 2016Updated 10 years ago
- Distributed t-SNE via Apache Spark☆158Dec 9, 2017Updated 8 years ago
- Zotero client build scripts☆11Apr 20, 2023Updated 3 years ago
- Generate images of neural network achitectures.☆20Sep 10, 2017Updated 9 years ago
- MLTK -- the Machine Learning Toolkit -- is a suite of C++ open source modules of Machine Learning.☆15Oct 25, 2013Updated 12 years ago
- High-performance key-value store☆12Dec 31, 2018Updated 7 years ago
- A Spark port of TFOCS: Templates for First-Order Conic Solvers (cvxr.com/tfocs)☆90Apr 15, 2024Updated 2 years ago
- Code to allow running BIDMach on Spark including HDFS integration and lightweight sparse model updates (Kylix).☆16Jul 23, 2020Updated 6 years ago
- writable stream for creating, appending, and replacing tables in cartodb☆11Nov 15, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Distributed Matrix Library☆73Jan 28, 2017Updated 9 years ago
- minio as local storage and DynamoDB as catalog☆15May 14, 2024Updated 2 years ago
- Simple role for deploying Elixir Exrm releases.☆10Jan 28, 2016Updated 10 years ago
- They only live to get radical.☆13Nov 29, 2018Updated 7 years ago
- Automatic offload of user-written Spark kernels to accelerators☆19Oct 25, 2016Updated 9 years ago
- Test pages listed in a sitemap.☆10Jan 7, 2015Updated 11 years ago
- Just-in-time Dynamic Batching with MXNet Gluon.☆52May 18, 2020Updated 6 years ago
- Blockchain data pipeline using Airflow, Kubernetes, Redshift, and Grafana☆18Mar 10, 2019Updated 7 years ago
- Adventures in robotics with Mindstorm EV3 and Elixir☆12Dec 30, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Scala library for neural networks☆107Mar 5, 2017Updated 9 years ago
- A support library for building hierarchies of composable 2nd-level schedulers. It uses parlib as its backend.☆17Jun 12, 2017Updated 9 years ago
- SBT template for projects written in Scala and other JVM languages☆13Dec 29, 2021Updated 4 years ago
- Code for "Adversarial Constraint Learning for Structured Prediction"☆14May 30, 2018Updated 8 years ago
- GM-PHD filter implementation in python (Gaussian mixture probability hypothesis density filter)☆10Nov 14, 2016Updated 9 years ago
- Kafka Connect Converter using JSONSchema☆14Oct 5, 2022Updated 3 years ago
- A tree-walk interpreter and a bytecode virtual machine interpreter written in the Rust Programming Language.☆13Jun 16, 2022Updated 4 years ago
- Javauto is a programming language for automation. Derived from Java, it is a cross platform alternative to something like AutoIt.☆16Jul 19, 2017Updated 9 years ago
- Application that visualizes your google location history in form of a heatmap using Spark to aggregate the data.☆12Feb 19, 2015Updated 11 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- PySpark for ETL jobs including lineage to Apache Atlas in one script via code inspection☆17Jan 12, 2017Updated 9 years ago
- an ansible role that installs elixir with kerl☆10Dec 14, 2018Updated 7 years ago
- A primal-dual framework for distributed L1-regularized optimization☆37Apr 18, 2016Updated 10 years ago
- The code for the in memory data pipeline that was presented at Berlin Buzzwords 2015.☆10Jun 1, 2015Updated 11 years ago
- A program that generates a cartoon and a caption using SVG paths and Markov chains☆10Nov 28, 2016Updated 9 years ago
- [In-Progress] Mini implementations of deep learning algorithms for natural language processing in PyTorch☆30Mar 30, 2017Updated 9 years ago
- Data science repo to help others☆12Feb 10, 2016Updated 10 years ago