Accelerator to rapidly deploy customized features for your business
☆57Dec 10, 2023Updated 2 years ago
Alternatives and similar repositories for feature-factory
Users that are interested in feature-factory are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Extensible Rules Engine for custom Dataframe / Dataset validation☆141May 7, 2024Updated 2 years ago
- Databricks Migration Tools☆43May 24, 2021Updated 5 years ago
- DEPRECATED: Integrating Jupyter with Databricks via SSH☆70Jun 28, 2022Updated 4 years ago
- Toolkit for Apache Spark ML for Feature clean-up, feature Importance calculation suite, Information Gain selection, Distributed SMOTE, Mo…☆191Jun 1, 2021Updated 5 years ago
- Manage your Databricks deployments and CI with code.☆203Feb 28, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Backend implementation for running MLFlow projects on Hadoop/YARN.☆11Dec 27, 2022Updated 3 years ago
- THIS PROJECT IS DEPRECATED. Capture deep metrics on one or all assets within a Databricks workspace☆229Jan 8, 2026Updated 9 months ago
- Generate big TPC-DS datasets with Databricks☆20Jan 3, 2022Updated 4 years ago
- A Chrome extension to apply a dark theme to Databricks notebooks☆15Dec 9, 2022Updated 3 years ago
- Example code for doing DataOps☆50Jan 26, 2021Updated 5 years ago
- This repo contains sample code and sample notebooks to illustrate how to work with Amazon FinSpace☆22Feb 12, 2025Updated last year
- Workshop for Spark and Databricks☆55Dec 6, 2019Updated 6 years ago
- Bulletproof Apache Spark jobs with fast root cause analysis of failures.☆73Mar 14, 2021Updated 5 years ago
- Scalefree's tool to automatically generate Data Vault models for the dbt package "datavault4dbt" based on metadata.☆50Oct 1, 2026Updated last week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code for Apache Hudi, Apache Iceberg and Delta Lake analysis☆10Feb 2, 2024Updated 2 years ago
- ☆10Jun 29, 2021Updated 5 years ago
- Data Exploration Using Spark 2.0☆14Apr 17, 2018Updated 8 years ago
- Code that was used as an example during the Data+AI Summit 2020☆15Mar 8, 2021Updated 5 years ago
- Code for P2V-MAP, a machine learning approach for mapping market structures☆11Dec 8, 2022Updated 3 years ago
- Dataset collected from popular Russian collective blog Habrahabr.ru☆12Oct 24, 2016Updated 9 years ago
- Unpack the source files from a Databricks .dbc archive file.☆26Mar 27, 2024Updated 2 years ago
- Building a real-time alert monitoring pipeline that sends email notifications off of Azure Event Hubs, Azure Databricks, and a Azure Logi…☆13Mar 8, 2020Updated 6 years ago
- The official Rock the JVM Akka Persistence Starter project☆11Apr 4, 2019Updated 7 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆31Mar 30, 2023Updated 3 years ago
- Script para importar dataset de "df_gtfs" a PostgreSQL☆14Jun 24, 2013Updated 13 years ago
- For Udemy students: the official repository for the Rock the JVM Akka HTTP with Scala course