Lightweight toolkit package to train and fine-tune 1.58bit Language models
☆149Aug 18, 2026Updated last month
Alternatives and similar repositories for onebitllms
Users that are interested in onebitllms are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Thin wrapper around GGML to make life easier☆48Jul 26, 2026Updated 2 months ago
- All information and news with respect to Falcon-H1 series☆127Oct 9, 2025Updated 11 months ago
- Rank-DistiLLM: Closing the Effectiveness Gap Between Cross-Encoders and LLMs for Passage Re-Ranking☆26Apr 4, 2025Updated last year
- YASEM - Yet Another Splade|Sparse Embedder - A simple and efficient library for SPLADE embeddings☆13May 22, 2025Updated last year
- An LLM Client for the PS Vita☆14Jun 23, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Evalution: evolve your LLMs with better evals.☆16Updated this week
- BitLinear implementation☆38Jul 8, 2026Updated 2 months ago
- PyTorch implementation of "Nextformer: A ConvNeXt Augmented Conformer For End-To-End Speech Recognition"☆10Dec 15, 2022Updated 3 years ago
- Starbucks: Improved Training for 2D Matryoshka Embeddings☆25Jun 30, 2025Updated last year
- Latent Large Language Models☆19Aug 24, 2024Updated 2 years ago
- Efficient non-uniform quantization with GPTQ for GGUF☆66Sep 17, 2025Updated last year
- Personal voice assistant, with voice interruption and Twilio support☆18Feb 24, 2025Updated last year
- Trully flash implementation of DeBERTa disentangled attention mechanism.☆93Sep 20, 2026Updated last week
- Linear Relational Embeddings (LREs) and Linear Relational Concepts (LRCs) for LLMs in PyTorch☆11Aug 7, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Minimalistic large language model 3D-parallelism training☆2,826Updated this week
- ☆54Jul 18, 2024Updated 2 years ago
- ☆16Dec 11, 2025Updated 9 months ago
- We release Open Meditron, a fully open, clinician-audited medical training corpus and evaluation protocol that closes the open-vs-closed …☆17Aug 3, 2026Updated last month
- Machine Bullshit: Characterizing the Emergent Disregard for Truth in Large Language Models☆26Sep 14, 2025Updated last year
- Pre-train Static Word Embeddings☆111Jun 9, 2026Updated 3 months ago
- Official implementation of Half-Quadratic Quantization (HQQ)☆960Feb 26, 2026Updated 7 months ago
- Official Repository for "Hypencoder: Hypernetworks for Information Retrieval"☆41Sep 20, 2025Updated last year
- Developing K - a language model to generate OPENSCAD code from prompt☆19Dec 3, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆19Sep 11, 2026Updated 2 weeks ago
- DPO, but faster 🚀☆51Dec 6, 2024Updated last year
- OpenAPI specifications => MCP (Model Context Protocol) tools☆19Dec 9, 2024Updated last year
- Python library to use Pleias-RAG models☆72Jul 1, 2026Updated 2 months ago
- This repository contains code for the MicroAdam paper.☆21Dec 14, 2024Updated last year
- An Open Source Toolkit For LLM Distillation☆1,069May 12, 2026Updated 4 months ago
- A powerful 130-million-parameter model trained from scratch as part of a truly open-source stack, including a custom tokenizer, dataset, …☆78Sep 2, 2025Updated last year
- chrome & firefox extension to chat with webpages: local llms☆132Dec 20, 2024Updated last year
- Running Microsoft's BitNet inference framework via FastAPI, Uvicorn and Docker.☆42Jul 2, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Repo hosting codes and materials related to speeding LLMs' inference using token merging.☆37Oct 9, 2025Updated 11 months ago
- A pipeline parallel training script for LLMs.☆169Apr 30, 2025Updated last year
- ☆64Jul 10, 2025Updated last year
- Estimate MFU for DeepSeekV3☆26Jan 5, 2025Updated last year
- EnriCo: Enriched Representation and Globally Constrained Inference for Entity and Relation Extraction☆27May 22, 2024Updated 2 years ago
- AI Edge Quantizer: flexible post training quantization for LiteRT models.☆206Updated this week
- Fast, Modern, and Low Precision PyTorch Optimizers☆130Sep 17, 2026Updated last week