Implementation of mamba with rust
☆96Mar 9, 2024Updated 2 years ago
Alternatives and similar repositories for mamba-ssm
Users that are interested in mamba-ssm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Inference of Mamba, Mamba2 and Mamba3 models in pure C☆202Mar 18, 2026Updated 4 months ago
- A single repo with all scripts and utils to train / fine-tune the Mamba model with or without FIM☆63Apr 8, 2024Updated 2 years ago
- win32 native frontend for llama-cli☆14Nov 2, 2024Updated last year
- Ultra-lightweight C++ inference engine for BitMamba-2 (1.58-bit SSM). Runs 1B models on consumer CPUs at 50+ tok/s using <700MB RAM. No h…☆21Jun 2, 2026Updated last month
- ☆20Jan 3, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Some preliminary explorations of Mamba's context scaling.☆221Feb 8, 2024Updated 2 years ago
- GPT-2 small trained on phi-like data☆68Feb 18, 2024Updated 2 years ago
- Implementation of the Mamba SSM with hf_integration.☆55Aug 31, 2024Updated last year
- ☆70Mar 1, 2024Updated 2 years ago
- A p2p patcher system written in rust using iroh and pkarr☆28Oct 26, 2025Updated 9 months ago
- ☆30Feb 27, 2024Updated 2 years ago
- BUD-E (Buddy) is an open-source voice assistant framework that facilitates seamless interaction with AI models and APIs, enabling the cre…☆44Jul 18, 2024Updated 2 years ago
- Modified Mamba code to run on CPU☆34Jan 14, 2024Updated 2 years ago
- Experience the power of AI with this free AI voice generator demo. Utilizing Deepgram and Groq, we transform text into voice seamlessly. …☆38Jun 12, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- My Implementation of Q-Sparse: All Large Language Models can be Fully Sparsely-Activated☆37Aug 14, 2024Updated last year
- Copy a bunch of files into your clipboard to provide context for LLMs☆115Feb 8, 2026Updated 5 months ago
- Exploring an idea where one forgets about efficiency and carries out attention across each edge of the nodes (tokens)☆56Mar 25, 2025Updated last year
- ☆27Feb 26, 2026Updated 4 months ago
- ☆12May 30, 2025Updated last year
- ☆16Oct 28, 2025Updated 8 months ago
- An efficient pytorch implementation of selective scan in one file, works with both cpu and gpu, with corresponding mathematical derivatio…☆109Oct 14, 2025Updated 9 months ago
- Evals meant to evaluate language models' ability to reason over long contexts.☆10Sep 12, 2024Updated last year
- FMS Model Optimizer is a framework for developing reduced precision neural network models.☆21Jun 24, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code repository for Black Mamba☆265Feb 8, 2024Updated 2 years ago
- OpenLine Protocol (OLP) — typed claim/evidence graphs for agentic systems with a 5-number digest. One-command FastAPI demo; auditable, re…☆15Feb 14, 2026Updated 5 months ago
- GPT-4 Level Conversational QA Trained In a Few Hours☆69Aug 21, 2024Updated last year
- Inference Llama 2 with a model compiled to native code by TorchInductor☆14Feb 8, 2024Updated 2 years ago
- AgentParse is a high-performance parsing library designed to map various structured data formats (such as Pydantic models, JSON, YAML, an…☆18Oct 13, 2025Updated 9 months ago
- A pure NumPy implementation of Mamba.☆221Jul 8, 2024Updated 2 years ago
- PyTorch implementation of "Nextformer: A ConvNeXt Augmented Conformer For End-To-End Speech Recognition"☆10Dec 15, 2022Updated 3 years ago
- Hardware-accelerated matrix/numeric programming library for Swift☆12Sep 2, 2025Updated 10 months ago
- ☆32Jan 7, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- LLM CLI Interface - Extremely Convenient and Fast☆12Sep 22, 2025Updated 10 months ago
- Simple, minimal implementation of the Mamba SSM in one file of PyTorch.☆2,965Mar 8, 2024Updated 2 years ago
- Inference Llama 2 in one file of pure Haskell (A port of llama2.c from Andrej Karpathy)☆14Oct 17, 2025Updated 9 months ago
- Mamba-Chat: A chat LLM based on the state-space model architecture 🐍☆943Mar 3, 2024Updated 2 years ago
- Lightweight Llama 3 8B Inference Engine in CUDA C☆53Mar 21, 2025Updated last year
- Implementation of a modular, high-performance, and simplistic mamba for high-speed applications☆41Nov 11, 2024Updated last year
- A preprint version of our recent research on the capability of frontier AI systems to do self-replication☆60Dec 18, 2024Updated last year