Lightweight toolkit package to train and fine-tune 1.58bit Language models
☆147Apr 30, 2026Updated 3 months ago
Alternatives and similar repositories for onebitllms
Users that are interested in onebitllms are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Thin wrapper around GGML to make life easier☆48Jul 26, 2026Updated 3 weeks ago
- All information and news with respect to Falcon-H1 series☆122Oct 9, 2025Updated 10 months ago
- Rank-DistiLLM: Closing the Effectiveness Gap Between Cross-Encoders and LLMs for Passage Re-Ranking☆25Apr 4, 2025Updated last year
- YASEM - Yet Another Splade|Sparse Embedder - A simple and efficient library for SPLADE embeddings☆13May 22, 2025Updated last year
- An LLM Client for the PS Vita☆14Jun 23, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Evalution: evolve your LLMs with better evals.☆16Aug 7, 2026Updated last week
- BitLinear implementation☆38Jul 8, 2026Updated last month
- PyTorch implementation of "Nextformer: A ConvNeXt Augmented Conformer For End-To-End Speech Recognition"☆10Dec 15, 2022Updated 3 years ago
- Starbucks: Improved Training for 2D Matryoshka Embeddings☆25Jun 30, 2025Updated last year
- Latent Large Language Models☆19Aug 24, 2024Updated last year
- Efficient non-uniform quantization with GPTQ for GGUF☆64Sep 17, 2025Updated 11 months ago
- Personal voice assistant, with voice interruption and Twilio support☆18Feb 24, 2025Updated last year
- Sparse Embedding Compression for Scalable Retrieval in Recommender Systems☆39Nov 21, 2025Updated 8 months ago
- Trully flash implementation of DeBERTa disentangled attention mechanism.☆91Feb 10, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Minimalistic large language model 3D-parallelism training☆2,785May 26, 2026Updated 2 months ago
- KSimply: An AI Potential Analyzer that recommends open-source models based on user hardware. / Un analizzatore di potenziale AI che consi…☆21Jun 2, 2026Updated 2 months ago
- ☆16Dec 11, 2025Updated 8 months ago
- Pre-train Static Word Embeddings☆110Jun 9, 2026Updated 2 months ago
- Official implementation of Half-Quadratic Quantization (HQQ)☆953Feb 26, 2026Updated 5 months ago
- Official Repository for "Hypencoder: Hypernetworks for Information Retrieval"☆41Sep 20, 2025Updated 10 months ago
- Developing K - a language model to generate OPENSCAD code from prompt☆19Dec 3, 2025Updated 8 months ago
- ☆18Oct 24, 2025Updated 9 months ago
- DPO, but faster 🚀☆51Dec 6, 2024Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Ultra-Sparse Adaptation of 1-Bit LLMs via XOR Patches☆86Aug 6, 2026Updated last week
- Python library to use Pleias-RAG models☆72Jul 1, 2026Updated last month
- An Open Source Toolkit For LLM Distillation☆1,018May 12, 2026Updated 3 months ago
- A powerful 130-million-parameter model trained from scratch as part of a truly open-source stack, including a custom tokenizer, dataset, …☆76Sep 2, 2025Updated 11 months ago
- Running Microsoft's BitNet inference framework via FastAPI, Uvicorn and Docker.☆40Jul 2, 2025Updated last year
- Repo hosting codes and materials related to speeding LLMs' inference using token merging.☆37Oct 9, 2025Updated 10 months ago
- A pipeline parallel training script for LLMs.☆167Apr 30, 2025Updated last year
- ☆10Aug 14, 2023Updated 3 years ago
- An efficent implementation of the method proposed in "The Era of 1-bit LLMs"☆155Oct 15, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The codebase and database of KuaiSearch: A Large-Scale E-Commerce Search Dataset for Recall, Ranking, and Relevance☆28Jul 27, 2026Updated 3 weeks ago
- AI Edge Quantizer: flexible post training quantization for LiteRT models.☆190Updated this week
- Estimate MFU for DeepSeekV3☆26Jan 5, 2025Updated last year
- EnriCo: Enriched Representation and Globally Constrained Inference for Entity and Relation Extraction☆26May 22, 2024Updated 2 years ago
- ☆63Jul 10, 2025Updated last year
- UniFork: Exploring Modality Alignment for Unified Multimodal Understanding and Generation☆48Aug 26, 2025Updated 11 months ago
- Fantasy Map Generation experiment on Python☆10Dec 9, 2020Updated 5 years ago