Insight Data Engineering project: A platform built in HDFS, Spark and Airflow to help you to find social influencers from GitHub Network.
☆16May 21, 2024Updated 2 years ago
Alternatives and similar repositories for Git-Influencer
Users that are interested in Git-Influencer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TrafficAdvisor: a Real-Time Traffic Monitoring System☆14Sep 10, 2018Updated 7 years ago
- Example project for consuming AWS Kinesis streamming and save data on Amazon Redshift using Apache Spark☆11May 22, 2018Updated 8 years ago
- A collection of data engineering projects: data modeling, ETL pipelines, data lakes, infrastructure configuration on AWS, data warehousin…☆15Apr 29, 2021Updated 5 years ago
- Project Search is a Recommendation system for Youtube videos and Amazon products.☆12May 10, 2017Updated 9 years ago
- Tweepy Stream Example☆19Apr 23, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Tool to query multiple metrics from a prometheus database through the REST API, and save them into a csv file.☆14May 8, 2018Updated 8 years ago
- Loan Default Prediction using PySpark, with jobs scheduled by Apache Airflow and Integration with Spark using Apache Livy☆22Dec 26, 2020Updated 5 years ago
- Usage examples for byte-genie API☆12Apr 27, 2024Updated 2 years ago
- This is a capstone project that entails building an end-to-end ETL (Extract-Transform-Load) Data pipeline which extracts UK accident and …☆18Jun 6, 2020Updated 5 years ago
- A production-grade data pipeline has been designed to automate the parsing of user search patterns to analyze user engagement. Extract d…☆24Nov 22, 2021Updated 4 years ago
- Internet's Most Popular Tutorials on Fresh-off-the-shelf ML & Data Science Technologies, Authored by Yours Truly.☆19Apr 29, 2020Updated 6 years ago
- Built a stream processing data pipeline to get data from disparate systems into a dashboard using Kafka as an intermediary.☆29Aug 14, 2023Updated 2 years ago
- Demonstration of using Apache Spark to build robust ETL pipelines while taking advantage of open source, general purpose cluster computin…☆24Aug 11, 2023Updated 2 years ago
- ☆19Feb 2, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A real-time event pipeline around Kafka Ecosystem for Chicago Transit Authority.☆32Aug 14, 2023Updated 2 years ago
- JobAnalytics system consumes data from multiple sources and provides valuable information to both job hunters and recruiters.☆30Dec 8, 2022Updated 3 years ago
- Spark data pipeline that processes movie ratings data.☆31May 1, 2026Updated 3 weeks ago
- Final and skeleton code for the clothing similarity walkthrough☆10Jan 20, 2016Updated 10 years ago
- Interactive Elasticsearch Analyzer☆13Dec 8, 2022Updated 3 years ago
- Project files for the post: Running PySpark Applications on Amazon EMR using Apache Airflow: Using the new Amazon Managed Workflows for A…☆41Jul 6, 2022Updated 3 years ago
- Analyzing shifting trends in music through the ages☆13Mar 25, 2021Updated 5 years ago
- general-purpose fast, stateless, and deterministic feature extractor written in golang for use in machine learning☆12Mar 17, 2018Updated 8 years ago
- Uploads files with background uploads and progress feedback on modern browsers☆10Mar 6, 2026Updated 2 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- the full stack☆13Jun 16, 2015Updated 10 years ago
- An implementation of the QUIC protocol in Elixir☆13Mar 17, 2019Updated 7 years ago
- Amazon Keyword Suggestion Tool in GoLang. Tool will generate relevant Amazon Product Keywords with the number of active products per each…☆50Jan 3, 2021Updated 5 years ago
- letter avatar is angular2 directive. It will generate avatar based on given text☆15Oct 31, 2019Updated 6 years ago
- A Yeoman generator for creating a FeathersJS plugin.☆22Aug 16, 2021Updated 4 years ago
- Pure Elixir implementation of Sha3 and the original Keccak1600-f☆16Jan 20, 2026Updated 4 months ago
- A collection of remark plugins used by HashiCorp to process markdown☆16Aug 22, 2025Updated 9 months ago
- epmd written in Elixir☆20Sep 24, 2014Updated 11 years ago
- Resources, notebooks, assets for ML for Everyone Twitch stream☆14Jul 8, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Annotate your pictures online and save in different formats☆14Oct 4, 2023Updated 2 years ago
- Learn how to use Kinesis Firehose, AWS Glue, S3, and Amazon Athena by streaming and analyzing reddit comments in realtime. 100-200 level …☆45Apr 20, 2021Updated 5 years ago
- Python writable in-memory virtual filesystem for SQLite☆17Jan 6, 2024Updated 2 years ago
- Jupyter notebook + Code for scraping AngelList data and making an interactive chart of SFBA salaries/equity☆14Jun 1, 2016Updated 9 years ago
- This is a version of Li Chen Wang's Palo Alto Tiny BASIC 2.0 for use with the online 8080 emulator and assembler ASM80.com.☆12Oct 10, 2020Updated 5 years ago
- ✋ Stop propagation for everyday events with Angular directives 🎩☆13Feb 4, 2018Updated 8 years ago
- Collection of Jupyter Notebooks in Python to monitor and improve your Watson Assistant workspaces☆10Jul 17, 2019Updated 6 years ago