Python package to accelerate the sparse matrix multiplication and top-n similarity selection
☆423Jun 8, 2026Updated this week
Alternatives and similar repositories for sparse_dot_topn
Users that are interested in sparse_dot_topn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Super Fast String Matching in Python☆370Jun 3, 2026Updated last week
- Group thousands of similar spreadsheet or database text entries in seconds☆158Jun 12, 2023Updated 2 years ago
- Spark Monitoring☆14Feb 28, 2023Updated 3 years ago
- Fuzzy string matching, grouping, and evaluation.☆798Jul 10, 2025Updated 11 months ago
- The privacy-preserving record linkage toolkit: a proof-of-concept public demo of next-gen data linkage techniques.☆16May 22, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ISO 20275☆10Oct 22, 2023Updated 2 years ago
- Monitor the stability of a Pandas or Spark dataframe ⚙︎☆512Jan 9, 2026Updated 5 months ago
- Abstractions for feature engineering on large graphs of tabular data.☆26May 18, 2026Updated 3 weeks ago
- Slack: #team-frontends-champions☆16Apr 24, 2025Updated last year
- Jupyter Widget to display resources used by the kernels☆13Aug 11, 2021Updated 4 years ago
- Google QUEST Q&A Labeling Kaggle Competition 6th Place Solution☆45Jun 11, 2020Updated 5 years ago
- Company Name Processor written in Python☆356Jan 16, 2026Updated 4 months ago
- Ordeq simplifies IO and modularizes pipeline logic.☆42Dec 19, 2025Updated 5 months ago
- Bag of, not words, but tricks!☆68Oct 31, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Rapid fuzzy string matching in Python using various string metrics☆3,946May 11, 2026Updated 3 weeks ago
- A powerful and modular toolkit for record linkage and duplicate detection in Python☆1,051Feb 21, 2024Updated 2 years ago
- Python wrapper for a C++ Double Metaphone☆15Jan 12, 2026Updated 4 months ago
- ☆13Dec 21, 2021Updated 4 years ago
- Entity Matching Model solves the problem of matching company names between two possibly very large datasets.☆95May 18, 2026Updated 3 weeks ago
- Python port of SymSpell: 1 million times faster spelling correction & fuzzy search through Symmetric Delete spelling correction algorithm…☆872Apr 20, 2026Updated last month
- Fast, accurate and scalable probabilistic data linkage with support for multiple SQL backends☆2,192Updated this week
- Set of tools to do parameter estimation from likelihood fits and estimate uncertainties on the fitted parameters or derived quantities.☆15May 22, 2019Updated 7 years ago
- Rich Context leaderboard competition, including the corpus and current SOTA for required tasks.☆22Nov 28, 2020Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- just a bunch of useful embeddings for scikit-learn pipelines☆526Feb 12, 2026Updated 3 months ago
- 📐 Compute distance between sequences. 30+ algorithms, pure python implementation, common interface, optional external libs usage.☆3,533Apr 18, 2025Updated last year
- Extra blocks for scikit-learn pipelines.☆1,397Updated this week
- 📛 Fuzzy Name Matching with Machine Learning☆268Jun 17, 2024Updated last year
- A python library for accurate and scalable fuzzy matching, record deduplication and entity-resolution.☆4,473Jul 29, 2025Updated 10 months ago
- Dataframe Integration with spaCy.☆103Mar 12, 2021Updated 5 years ago
- Doubt your data, find bad labels.☆516Jul 15, 2024Updated last year
- Python package for Model Metric Uncertainty estimation☆17Updated this week
- Approximate Nearest Neighbor Search for Sparse Data in Python!☆918Oct 2, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- MinHash, LSH, LSH Forest, Weighted MinHash, HyperLogLog, HyperLogLog++, LSH Ensemble and HNSW☆2,928Apr 18, 2026Updated last month
- Toolkit to help understand "what lies" in word embeddings. Also benchmarking!☆480Feb 6, 2023Updated 3 years ago
- ☆32Dec 15, 2023Updated 2 years ago
- Light-weight, Python-based data-analysis framework☆12Feb 24, 2019Updated 7 years ago
- Tree-based indexes for neural-search☆33Mar 4, 2024Updated 2 years ago
- kagglerが使いそうなslack emojiをまとめたリポジトリだよ。☆21Feb 20, 2022Updated 4 years ago
- Efficient Trie-based regex unions for blacklist/whitelist filtering and one-pass mapping-based string replacing☆76May 1, 2026Updated last month