☆55Feb 10, 2025Updated last year
Alternatives and similar repositories for ModernBERT-Instruct-mini-cookbook
Users that are interested in ModernBERT-Instruct-mini-cookbook are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ModernVBERT is a 250M-parameter vision–language encoder that aligns a text-encoder (Ettin-150M) with a vision-encoder (SigLIP2-B) through…☆16Oct 16, 2025Updated 11 months ago
- ☆109Jun 2, 2025Updated last year
- ☆14Oct 21, 2024Updated last year
- Luth is a state-of-the-art series of fine-tuned LLMs for French☆47Oct 12, 2025Updated 11 months ago
- Efficient and modular GraphRAG system☆50Jul 7, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Benchmark harness for evaluating DSPy RLMs on data analysis tasks (InfiAgent-DABench)☆24Mar 22, 2026Updated 6 months ago
- Generalist and Lightweight Model for Text Classification☆504Jul 21, 2026Updated 2 months ago
- Transform is the main building block of data pipelines in fastai. And elsewhere if you want.☆35Updated this week
- recipe for training fully-featured self supervised image jepa models☆14Jun 4, 2025Updated last year
- ☆37May 5, 2025Updated last year
- code for training & evaluating Contextual Document Embedding models☆207May 14, 2025Updated last year
- Code for SaGe subword tokenizer (EACL 2023)☆28Nov 30, 2024Updated last year
- Official implementation of "GPT or BERT: why not both?"☆65Jul 28, 2025Updated last year
- Materials for the "My Workflow for Understanding LLM Architectures" tutorial☆30Apr 10, 2026Updated 5 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This repository contains the training and evaluation code for llm-jp-modernbert-base.☆17Jun 17, 2025Updated last year
- Label shift estimation for transfer difficulty with Familiarity.☆10Feb 4, 2025Updated last year
- Late Interaction Models Training & Retrieval☆895Jul 23, 2026Updated last month
- Python library to use Pleias-RAG models☆72Jul 1, 2026Updated 2 months ago
- ☆104Jul 4, 2025Updated last year
- Jax like function transformation engine but micro, microjax☆34Oct 25, 2024Updated last year
- ☆59Aug 19, 2025Updated last year
- ☆10Oct 22, 2024Updated last year
- A data visualisation of a 100 responses when asking local LLMs to imagine a random person.☆24Nov 4, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Bringing BERT into modernity via both architecture changes and scaling☆1,738Mar 1, 2026Updated 6 months ago
- This repository contains papers for a comprehensive survey on accelerated generation techniques in Large Language Models (LLMs).☆11May 24, 2024Updated 2 years ago
- Example dialogs to get your creative juices flowing☆46Mar 31, 2026Updated 5 months ago
- Code for the Avey-B paper (https://arxiv.org/abs/2602.15814)☆32Feb 21, 2026Updated 7 months ago
- Efficient and scalable zero-shot entity linking☆155Jul 20, 2026Updated 2 months ago
- A context-aware embedding similarity score☆11Aug 23, 2023Updated 3 years ago
- Examples for using the SiLLM framework for training and running Large Language Models (LLMs) on Apple Silicon☆16May 8, 2025Updated last year
- This is a question-output workflow template for shiny app!☆12May 17, 2019Updated 7 years ago
- ☆67Mar 4, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [ICML2026] Official JAX code for Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying☆17Jul 3, 2026Updated 2 months ago
- [ICML 2024] Learning with Complementary Labels Revisited: The Selected-Completely-at-Random Setting Is More Practical☆13May 12, 2024Updated 2 years ago
- ☆29Aug 21, 2025Updated last year
- PySpark custom data source for Hugging Face Datasets☆27Updated this week
- Efficient encoder-decoder architecture for small language models (≤1B parameters) with cross-architecture knowledge distillation and visi…☆33Feb 7, 2025Updated last year
- LexiSignVQA: A Unified Training-free Multi-stage Approach to Multimodal Legal Question Answering on Traffic Sign Rules☆23Nov 18, 2025Updated 10 months ago
- ☆13Dec 17, 2021Updated 4 years ago