Implementation of mamba with rust
☆97Mar 9, 2024Updated 2 years ago
Alternatives and similar repositories for mamba-ssm
Users that are interested in mamba-ssm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Inference of Mamba, Mamba2 and Mamba3 models in pure C☆203Mar 18, 2026Updated 5 months ago
- A single repo with all scripts and utils to train / fine-tune the Mamba model with or without FIM☆63Apr 8, 2024Updated 2 years ago
- win32 native frontend for llama-cli☆14Nov 2, 2024Updated last year
- Official Repository for Efficient Linear-Time Attention Transformers.☆17Jun 2, 2024Updated 2 years ago
- A rust wrapper for the spoa C++ partial order alignment library☆10Jun 11, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- GPT-2 small trained on phi-like data☆68Feb 18, 2024Updated 2 years ago
- Implementation of the Mamba SSM with hf_integration.☆56Aug 31, 2024Updated 2 years ago
- PanGenome Graph Building with the first 100 assemblies from the 1000G ONT Sequencing Consortium☆14Apr 5, 2025Updated last year
- ☆30Feb 27, 2024Updated 2 years ago
- Yet another frontend for LLM, written using .NET and WinUI 3☆11Sep 14, 2025Updated 11 months ago
- BUD-E (Buddy) is an open-source voice assistant framework that facilitates seamless interaction with AI models and APIs, enabling the cre…☆43Jul 18, 2024Updated 2 years ago
- Modified Mamba code to run on CPU☆34Jan 14, 2024Updated 2 years ago
- Experience the power of AI with this free AI voice generator demo. Utilizing Deepgram and Groq, we transform text into voice seamlessly. …☆38Jun 12, 2024Updated 2 years ago
- My Implementation of Q-Sparse: All Large Language Models can be Fully Sparsely-Activated☆37Aug 14, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Rust binding for WFA2-lib☆10Jun 7, 2022Updated 4 years ago
- Lightweight C inference for Qwen3 GGUF. Multiturn prefix caching & batch processing.☆26Sep 1, 2025Updated last year
- an auto-sleeping and -waking framework around llama.cpp☆13Feb 8, 2025Updated last year
- Exploring an idea where one forgets about efficiency and carries out attention across each edge of the nodes (tokens)☆56Mar 25, 2025Updated last year
- ☆28Feb 26, 2026Updated 6 months ago
- ☆12May 30, 2025Updated last year
- ☆16Oct 28, 2025Updated 10 months ago
- cortex.llamacpp is a high-efficiency C++ inference engine for edge computing. It is a dynamic library that can be loaded by any server a…☆44Jul 4, 2025Updated last year
- Group-relative Trajectory-based Policy Optimization: Increasing Quality and Training Stability☆42Feb 23, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- An efficient pytorch implementation of selective scan in one file, works with both cpu and gpu, with corresponding mathematical derivatio…☆109Oct 14, 2025Updated 10 months ago
- Evals meant to evaluate language models' ability to reason over long contexts.☆10Sep 12, 2024Updated last year
- FMS Model Optimizer is a framework for developing reduced precision neural network models.☆21Jun 24, 2026Updated 2 months ago
- Code repository for Black Mamba☆265Feb 8, 2024Updated 2 years ago
- ☆12Apr 4, 2024Updated 2 years ago
- 更纯粹、更高压缩率的Tokenizer in Rust☆14Dec 21, 2024Updated last year
- GPT-4 Level Conversational QA Trained In a Few Hours☆69Aug 21, 2024Updated 2 years ago
- Inference Llama 2 with a model compiled to native code by TorchInductor☆14Feb 8, 2024Updated 2 years ago
- NLSpec instruction following benchmark for https://factory.strongdm.ai/products/attractor☆21Feb 26, 2026Updated 6 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- AgentParse is a high-performance parsing library designed to map various structured data formats (such as Pydantic models, JSON, YAML, an…☆18Oct 13, 2025Updated 10 months ago
- Load and run Llama from safetensors files in C☆16Oct 24, 2024Updated last year
- A pure NumPy implementation of Mamba.☆221Jul 8, 2024Updated 2 years ago
- ☆32Jan 7, 2024Updated 2 years ago
- Implementation for <Robust Weight Perturbation for Adversarial Training> in IJCAI'22.☆16Jul 1, 2022Updated 4 years ago
- Simple, minimal implementation of the Mamba SSM in one file of PyTorch.☆2,967Mar 8, 2024Updated 2 years ago
- Inference Llama 2 in one file of pure Haskell (A port of llama2.c from Andrej Karpathy)☆14Oct 17, 2025Updated 10 months ago