Implementation of mamba with rust
☆97Mar 9, 2024Updated 2 years ago
Alternatives and similar repositories for mamba-ssm
Users that are interested in mamba-ssm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Inference of Mamba, Mamba2 and Mamba3 models in pure C☆203Mar 18, 2026Updated 4 months ago
- A single repo with all scripts and utils to train / fine-tune the Mamba model with or without FIM☆63Apr 8, 2024Updated 2 years ago
- win32 native frontend for llama-cli☆14Nov 2, 2024Updated last year
- Official Repository for Efficient Linear-Time Attention Transformers.☆17Jun 2, 2024Updated 2 years ago
- Production-ready ternary quantized (1.58-bit) Rust code generation model with mHC-lite, MaxRL training, and comprehensive benchmarking☆22Aug 2, 2026Updated 2 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Ultra-lightweight C++ inference engine for BitMamba-2 (1.58-bit SSM). Runs 1B models on consumer CPUs at 50+ tok/s using <700MB RAM. No h…☆20Jun 2, 2026Updated 2 months ago
- A rust wrapper for the spoa C++ partial order alignment library☆10Jun 11, 2025Updated last year
- ☆20Jan 3, 2024Updated 2 years ago
- GPT-2 small trained on phi-like data☆68Feb 18, 2024Updated 2 years ago
- Implementation of the Mamba SSM with hf_integration.☆56Aug 31, 2024Updated last year
- ☆70Mar 1, 2024Updated 2 years ago
- A p2p patcher system written in rust using iroh and pkarr☆28Oct 26, 2025Updated 9 months ago
- PanGenome Graph Building with the first 100 assemblies from the 1000G ONT Sequencing Consortium☆14Apr 5, 2025Updated last year
- ☆30Feb 27, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Yet another frontend for LLM, written using .NET and WinUI 3☆11Sep 14, 2025Updated 11 months ago
- Modified Mamba code to run on CPU☆34Jan 14, 2024Updated 2 years ago
- Experience the power of AI with this free AI voice generator demo. Utilizing Deepgram and Groq, we transform text into voice seamlessly. …☆38Jun 12, 2024Updated 2 years ago
- My Implementation of Q-Sparse: All Large Language Models can be Fully Sparsely-Activated☆37Aug 14, 2024Updated 2 years ago
- Copy a bunch of files into your clipboard to provide context for LLMs☆115Feb 8, 2026Updated 6 months ago
- It is almost the best 3B model in the current open source industry, surpassing Dolly v2-3b, open lama-3b, and even outperforming the Eleu…☆15Jul 24, 2023Updated 3 years ago
- Lightweight C inference for Qwen3 GGUF. Multiturn prefix caching & batch processing.☆25Sep 1, 2025Updated 11 months ago
- an auto-sleeping and -waking framework around llama.cpp☆13Feb 8, 2025Updated last year
- Exploring an idea where one forgets about efficiency and carries out attention across each edge of the nodes (tokens)☆56Mar 25, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆27Feb 26, 2026Updated 5 months ago
- ☆12May 30, 2025Updated last year
- Group-relative Trajectory-based Policy Optimization: Increasing Quality and Training Stability☆42Feb 23, 2026Updated 5 months ago
- An efficient pytorch implementation of selective scan in one file, works with both cpu and gpu, with corresponding mathematical derivatio…☆109Oct 14, 2025Updated 10 months ago
- Evals meant to evaluate language models' ability to reason over long contexts.☆10Sep 12, 2024Updated last year
- FMS Model Optimizer is a framework for developing reduced precision neural network models.☆21Jun 24, 2026Updated last month
- Code repository for Black Mamba☆265Feb 8, 2024Updated 2 years ago
- ☆12Apr 4, 2024Updated 2 years ago
- 更纯粹、更高压缩率的Tokenizer in Rust☆14Dec 21, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- GPT-4 Level Conversational QA Trained In a Few Hours☆69Aug 21, 2024Updated last year
- Inference Llama 2 with a model compiled to native code by TorchInductor☆14Feb 8, 2024Updated 2 years ago
- AgentParse is a high-performance parsing library designed to map various structured data formats (such as Pydantic models, JSON, YAML, an…☆18Oct 13, 2025Updated 10 months ago
- Load and run Llama from safetensors files in C☆16Oct 24, 2024Updated last year
- A pure NumPy implementation of Mamba.☆221Jul 8, 2024Updated 2 years ago
- ☆32Jan 7, 2024Updated 2 years ago
- LLM CLI Interface - Extremely Convenient and Fast☆12Sep 22, 2025Updated 10 months ago