A simple implementation of [Mamba: Linear-Time Sequence Modeling with Selective State Spaces](https://arxiv.org/abs/2312.00752)
☆23Jan 22, 2024Updated 2 years ago
Alternatives and similar repositories for Mamba_SSM
Users that are interested in Mamba_SSM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15May 8, 2017Updated 9 years ago
- ☆24Sep 25, 2024Updated last year
- ☆29Jan 17, 2025Updated last year
- The demo projects for Allwinner D1 SBC☆24Sep 7, 2021Updated 4 years ago
- Here we will test various linear attention designs.☆62Apr 25, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Implementation of a Hierarchical Mamba as described in the paper: "Hierarchical State Space Models for Continuous Sequence-to-Sequence Mo…☆16Nov 11, 2024Updated last year
- AlphaZero implementation on Gomoku☆18Feb 26, 2025Updated last year
- AnyDSL traversal code☆15Feb 18, 2019Updated 7 years ago
- Distributed machine learning platform☆13Aug 20, 2015Updated 10 years ago
- OpenEarthMap-SAR: A benchmark dataset for land cover mapping under all-weather conditions☆20Jun 26, 2025Updated last year
- Complete solution to enable RDMA (on both InfiniBand and RoCE) and accelerate TCP to bare metal performance on Kubernetes☆11Aug 1, 2018Updated 7 years ago
- A PyTorch implementation of "Self-Supervised GNN that Jointly Learns to Augment" or "Jointly Learnable Data Augmentations for Self-Superv…☆13Dec 13, 2021Updated 4 years ago
- ☆16Jul 24, 2023Updated 3 years ago
- Subpart source code of of deepcore v0.7☆27Jun 28, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- POPGym Library in JAX☆14Apr 15, 2024Updated 2 years ago
- A fusion of a linear layer and a cross entropy loss, written for pytorch in triton.☆75Aug 2, 2024Updated last year
- Source code for the paper "LongGenBench: Long-context Generation Benchmark"☆24Oct 8, 2024Updated last year
- DocReal: Robust Document Dewarping of Real-Life Images via Attention-Enhanced Control Point Prediction☆30Jun 28, 2023Updated 3 years ago
- Accelerated First Order Parallel Associative Scan☆198Jan 7, 2026Updated 6 months ago
- Implementation of MambaByte in "MambaByte: Token-free Selective State Space Model" in Pytorch and Zeta☆128Jul 20, 2026Updated last week
- BLAS OpenCL implementation.☆17Apr 8, 2015Updated 11 years ago
- PyTorch implementation of Continuously Indexed Flows paper, with many baseline normalising flows☆32Sep 16, 2021Updated 4 years ago
- ☆13Dec 18, 2020Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆13Jul 3, 2024Updated 2 years ago
- DeepPerf is a set of cuda assembling developing tools☆11Dec 19, 2018Updated 7 years ago
- [ICCV 2021] Code release for "Sub-bit Neural Networks: Learning to Compress and Accelerate Binary Neural Networks"☆33Jul 24, 2022Updated 4 years ago
- Scripts and related files to install Nik for GIMP on Ubuntu☆17May 26, 2025Updated last year
- Jax implementation of "Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models"☆15May 10, 2024Updated 2 years ago
- ☆29Jul 9, 2024Updated 2 years ago
- This repo contains the official implementation of paper "Layered Controllabel Video Generation".☆13Oct 31, 2022Updated 3 years ago
- Application of Higher-Order Singular Value Decomposition (HOSVD) for a microseismic dataset.☆14Jun 11, 2021Updated 5 years ago
- ICME2022 Special Session “Beyond Accuracy: Responsible, Responsive, and Robust Multimedia Retrieval ”☆12Jun 3, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆17Aug 22, 2021Updated 4 years ago
- Variational Reinforcement Learning☆18Jul 25, 2024Updated 2 years ago
- MurmurHash3 x64 128-bit - a fast, non-cryptographic hash function☆22Jun 2, 2020Updated 6 years ago
- ☆107Mar 9, 2024Updated 2 years ago
- Vocoder-Free Non-Parallel Conversion of Whispered Speech With Masked Cycle-Consistent Generative Adversarial Networks☆17Aug 18, 2023Updated 2 years ago
- LLM Inference with Microscaling Format☆35Nov 12, 2024Updated last year
- benchmark models for TNN, ncnn, MNN☆21Jun 10, 2020Updated 6 years ago