☆58Sep 26, 2025Updated 11 months ago
Alternatives and similar repositories for infini-gram-mini
Users that are interested in infini-gram-mini are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆114Jan 24, 2026Updated 7 months ago
- ☆105Jul 16, 2026Updated last month
- Toolkit for domain-specific information retrieval experimentation☆19May 18, 2026Updated 3 months ago
- The official repository for the paper entitled "Time Travel in LLMs: Tracing Data Contamination in Large Language Models."☆14Jun 11, 2024Updated 2 years ago
- SSLCL: An Efficient Model-Agnostic Supervised Contrastive Learning Framework for Emotion Recognition in Conversations☆15Jul 27, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Efficiently computing & storing token n-grams from large corpora☆28Jun 15, 2026Updated 2 months ago
- Joint Selection for Large-Scale Pre-Training Data via Policy Gradient-based Mask Learning☆21Jan 4, 2026Updated 7 months ago
- A Jensen-Shannon Divergence Driven Mechanistic Study of Context Attribution in Retrieval-Augmented Generation☆16Aug 28, 2025Updated last year
- Ukrainian ELECTRA model☆12Mar 11, 2023Updated 3 years ago
- ☆13Aug 20, 2021Updated 5 years ago
- ☆22Dec 1, 2021Updated 4 years ago
- The source code for the TIRA Shared Task Platform☆19Updated this week
- Efficient Language Model Training through Cross-Lingual and Progressive Transfer Learning☆30Jan 25, 2023Updated 3 years ago
- [ACL2025 Best Paper] Language Models Resist Alignment☆52Jun 11, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A library for language transfer methods and algorithms.☆16Feb 6, 2026Updated 6 months ago
- [COLM 2024] Early Weight Averaging meets High Learning Rates for LLM Pre-training☆19Oct 12, 2024Updated last year
- ☆10Mar 1, 2025Updated last year
- suffix array construction and searching algorithms for in-memory binary data.☆13Sep 10, 2022Updated 3 years ago
- From Hero to Zéroe: A Benchmark of Low-Level Adversarial Attacks☆15Feb 23, 2023Updated 3 years ago
- decontamination☆38Mar 4, 2026Updated 5 months ago
- LIGHTVOC AN UPSAMPLING-FREE GAN VOCODER BASED ON CONFORMER AND INVERSE SHORT-TIME FOURIER TRANSFORM☆18May 17, 2024Updated 2 years ago
- ☆18Oct 11, 2025Updated 10 months ago
- [EMNLP'23] Official Code for "FOCUS: Effective Embedding Initialization for Monolingual Specialization of Multilingual Models"☆37Jun 7, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Python module to remove wiki markup text.☆10Jan 15, 2016Updated 10 years ago
- ☆22Dec 11, 2024Updated last year
- ☆52Jan 24, 2024Updated 2 years ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated last year
- ☆16Jan 30, 2025Updated last year
- Code for NeurIPS 2024 Paper - Superposed Decoding: Multiple Generations from a Single Autoregressive Inference Pass☆21Aug 22, 2024Updated 2 years ago
- Deep learning approaches in detecting 14 different abnormalities via Chest X-Ray images☆11Jan 16, 2022Updated 4 years ago
- Lean4 Code Editor☆18Aug 18, 2026Updated last week
- ☆25Dec 9, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code for Bolmo: Byteifying the Next Generation of Language Models☆136Updated this week
- Code and data setup for the paper "Are Diffusion Models Vision-and-language Reasoners?"☆33Mar 15, 2024Updated 2 years ago
- [EMNLP 2025 Main] ConceptVectors Benchmark and Code for the paper "Intrinsic Evaluation of Unlearning Using Parametric Knowledge Traces"☆39Aug 20, 2025Updated last year
- Code for ICML 2025 paper | Joint Localization and Activation Editing for Low-Resource Fine-Tuning☆28Jun 18, 2025Updated last year
- Code for WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models.☆93Sep 12, 2024Updated last year
- Forcing Diffuse Distributions out of Language Models☆18Sep 10, 2024Updated last year
- An unbounded n-gram language model on Tiny Shakespeare☆22Jan 21, 2026Updated 7 months ago