mamba2-jax: A pure JAX/Flax implementation of Mamba-2 for language modeling and time series forecasting.
☆19Jun 23, 2026Updated 3 months ago
Alternatives and similar repositories for mamba2-jax
Users that are interested in mamba2-jax are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An open source reinforcement learning codebase with a variety of intrinsic exploration methods implemented in PyTorch.☆11Feb 6, 2023Updated 3 years ago
- Amazon Alexa Voice Controlled Drone☆10Jan 18, 2020Updated 6 years ago
- Ukrainian ELECTRA model☆12Mar 11, 2023Updated 3 years ago
- This is a code repository for Relation Transformer Network☆13Nov 30, 2021Updated 4 years ago
- Code for reproducing the paper "Neural Networks Fail to Learn Periodic Functions and How to Fix It" as part of the ML Reproducibility Cha…☆11Apr 16, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- suffix array construction and searching algorithms for in-memory binary data.☆13Sep 10, 2022Updated 4 years ago
- From Hero to Zéroe: A Benchmark of Low-Level Adversarial Attacks☆15Feb 23, 2023Updated 3 years ago
- Official implementation for the CVPR 2024 paper: HuProSO3: Normalizing Flows on the Product Space of SO(3) Manifolds for Probabilistic Hu…☆17Mar 31, 2025Updated last year
- Python module to remove wiki markup text.☆10Jan 15, 2016Updated 10 years ago
- Network representation learning on drug-target-side effects-indication graphs for side effect prediction☆13Feb 4, 2020Updated 6 years ago
- a minimal website to get the diff of llm rewrites☆11Dec 11, 2024Updated last year
- The official code of TACL 2022, "Break, Perturb, Build: Automatic Perturbation of Reasoning Paths Through Question Decomposition".☆12Oct 18, 2021Updated 4 years ago
- Following research on S4 in jax☆16Jun 15, 2022Updated 4 years ago
- High quality implementations of imitation and inverse reinforcement learning algorithms☆26Aug 19, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Fast reinforcement learning 💨☆29Jul 15, 2025Updated last year
- ☆10Jul 14, 2018Updated 8 years ago
- Code for the paper "BPE stays on SCRIPT", "Which Pieces Does Unigram Tokenization Really Need?" and MinGram☆22Sep 22, 2026Updated 2 weeks ago
- ☆10Jan 28, 2019Updated 7 years ago
- High-performance backend for language model probabilistic programs☆17Sep 26, 2026Updated last week
- Cheat sheet for the "Deep Learning" course at ETH Zürich☆22Nov 13, 2019Updated 6 years ago
- GCRL in JAX. Official repository for LEO (ICML 2026).☆34Jun 20, 2026Updated 3 months ago
- Repository for "Self-Distillation for Model Stacking Unlocks Cross-Lingual NLU in 200+ Languages"☆15Oct 4, 2024Updated 2 years ago
- ☆18Apr 24, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Archives for Triton Inference Server Practices☆15Feb 28, 2022Updated 4 years ago
- Python package for Geometric / Clifford Algebra with Pytorch.☆16Jun 2, 2026Updated 4 months ago
- ☆23Aug 21, 2021Updated 5 years ago
- Repository for Sparse Universal Transformers☆20Oct 23, 2023Updated 2 years ago
- (SIGGRAPH Asia 2025) Pytorch implementation of "TrackerSplat: Exploiting Point Tracking for Fast and Robust Dynamic 3D Gaussians Reconst…☆21Apr 26, 2026Updated 5 months ago
- Scaling Sparse Fine-Tuning to Large Language Models☆20Jan 31, 2024Updated 2 years ago
- Code for the paper "Getting the most out of your tokenizer for pre-training and domain adaptation"☆22Feb 14, 2024Updated 2 years ago
- Developing hybrid deep learning models by integrating Neural networks with (s,e,t)GARCH models to predict volatility in the Indian Commod…☆19May 21, 2021Updated 5 years ago
- AI powered Virtual Desktop☆18Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [RSS 2026] LPS: Latent Policy Steering through One-Step Flow Policies.☆20Sep 18, 2026Updated 3 weeks ago
- Diversity is All You Need: Learning Skills without a Reward Function in PyTorch.☆90Jan 12, 2026Updated 8 months ago
- An extention to the GaLore paper, to perform Natural Gradient Descent in low rank subspace☆19Oct 21, 2024Updated last year
- An arbitrage bot is a smart contract connected to an external automation script that controls its operation.☆2,789Updated this week
- streaming deep reinforcement learning but 4x faster with jax!☆19Jan 4, 2026Updated 9 months ago
- ☆17Aug 20, 2025Updated last year
- <알파제로를 분석하며 배우는 인공지능> 리포지토리☆14Feb 18, 2020Updated 6 years ago