Compression for unit-norm embedding vectors using spherical coordinates
☆83Jan 23, 2026Updated 5 months ago
Alternatives and similar repositories for jzip-compressor
Users that are interested in jzip-compressor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The training codes of Jasper-Token-Compression-600M☆20Nov 19, 2025Updated 8 months ago
- PyLate efficient inference engine☆87Jan 7, 2026Updated 6 months ago
- Generate fixed dimensional embeddings for multi-dimensional vectors in python based on Muvera from Google.☆20Jun 28, 2025Updated last year
- ☆27Jun 11, 2026Updated last month
- Datamodels for hugging face tokenizers☆108Jun 18, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICML'26] LEMUR reduces multi-vector retrieval for late interaction models such as ColBERT into regular single-vector retrieval.☆31Jun 21, 2026Updated last month
- Late Interaction Models Training & Retrieval☆875Jul 13, 2026Updated last week
- Cuda implemenation of flash-kmeans, 2x faster☆23May 8, 2026Updated 2 months ago
- A framework for benchmarking embedding models in hybrid search scenarios (BM25 + vector search) using Weaviate.☆40Updated this week
- An extensive and commented list of resources on Late-Interaction Multivector Retrieval.☆67Jul 8, 2026Updated last week
- No code tool for finetuning embedding models☆30Updated this week
- Fast search index for SPLADE sparse retrieval models implemented in Python using Numpy and Numba☆38Oct 16, 2025Updated 9 months ago
- HSEB: Hybrid Search Engine Benchmark☆21Oct 5, 2025Updated 9 months ago
- A Rust rewrite of FastKMeans for CPU-based clustering☆17Jun 29, 2026Updated 3 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Can VLMs understand students' hand-drawn math work?☆19Jan 20, 2026Updated 6 months ago
- Crispy reranking models by Mixedbread☆52Sep 17, 2025Updated 10 months ago
- High-Performance Engine for Multi-Vector Search☆268May 28, 2026Updated last month
- Examples for the Activate conference☆11Sep 11, 2019Updated 6 years ago
- Latent Large Language Models☆19Aug 24, 2024Updated last year
- Official implementation of "MMNeuron: Discovering Neuron-Level Domain-Specific Interpretation in Multimodal Large Language Model". Our co…☆26Dec 20, 2024Updated last year
- Production inference for encoder models - ColBERT, GLiNER, ColPali, embeddings etc. - as vLLM plugins for online and in-process deploymen…☆75Jul 6, 2026Updated 2 weeks ago
- PromptMII: Meta-Learning Instruction Induction for LLMs☆48Jan 12, 2026Updated 6 months ago
- utilities for loading and running text embeddings with onnx☆46Aug 16, 2025Updated 11 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Neo4J Plugin for Protege , Import export ontologies from Neo4j into protege, write queries in NLP☆15Dec 11, 2025Updated 7 months ago
- ☆81Dec 12, 2025Updated 7 months ago
- Official repository of TACHIOM.☆62Updated this week
- An extensive and commented list of resources on Learned Sparse Retrieval.☆63Updated this week
- Nearly Inference Free Embeddings: make your RAG queries 500x faster☆80Apr 27, 2026Updated 2 months ago
- Faster Learned Sparse Retrieval with Block-Max Pruning. ACM SIGIR 2024.☆37Jan 14, 2026Updated 6 months ago
- LLM-as-a-judge using G-eval Scratch☆15Oct 12, 2025Updated 9 months ago
- Tree-based indexes for neural-search☆33Mar 4, 2024Updated 2 years ago
- ⚡ Super fast clustering for high-dimensional vectors on CPUs (x86, ARM) and GPUs — for Python and C++. 100x faster clustering of vector e…☆67Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Zero and Few-shot document level relation extraction / ⚠️ Development moved to: https://github.com/cea-list-lasti/glidre☆18Mar 13, 2026Updated 4 months ago
- Landing repository for the paper "Predicting the Order of Upcoming Tokens Improves Language Modeling"☆48May 13, 2026Updated 2 months ago
- unofficial implementation of MUVERA: Multi-Vector Retrieval via Fixed Dimensional Encodings☆15Feb 18, 2026Updated 5 months ago
- Storing the LongCoT-mini results for RLM(GPT-5.2)☆20Apr 26, 2026Updated 2 months ago
- Official repository of the Seismic library.☆135Jul 6, 2026Updated 2 weeks ago
- ☆23May 30, 2025Updated last year
- Embedding Inversion via Conditional Masked Diffusion: recover original text from embedding vectors using parallel denoising. Live demo + …☆60Mar 7, 2026Updated 4 months ago