ROSA+: RWKV's ROSA implementation with fallback statistical predictor
☆36Oct 13, 2025Updated 9 months ago
Alternatives and similar repositories for rosa-plus
Users that are interested in rosa-plus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Softened ROSA QKV Operators for Training Next-Generation LLM Models☆39Jun 26, 2026Updated 3 weeks ago
- ROSA-Tuning☆74Feb 4, 2026Updated 5 months ago
- RWKV-7 7.2B fp16 15000+ tps decoding @ single 5090☆122Updated this week
- RADLADS training code☆46May 7, 2025Updated last year
- ☆48Jul 3, 2026Updated 3 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- https://x.com/BlinkDL_AI/status/1884768989743882276☆28May 4, 2025Updated last year
- A high-efficiency text embedding and reranking model based on RWKV architecture.☆19Jan 10, 2026Updated 6 months ago
- ☆41Apr 30, 2025Updated last year
- ☆13Aug 19, 2024Updated last year
- Efficient implementations of state-of-the-art linear attention models in Pytorch and Triton☆50Apr 2, 2026Updated 3 months ago
- continous batching and parallel acceleration for RWKV6☆23Jun 28, 2024Updated 2 years ago
- RWKV-LM-V7(https://github.com/BlinkDL/RWKV-LM) Under Lightning Framework☆62May 13, 2026Updated 2 months ago
- MLX binary vectors and associated algorithms.☆14Mar 13, 2025Updated last year
- This repo is an exploratory experiment to enable frozen pretrained RWKV language models to accept speech modality input. We followed the …☆54Dec 23, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A port of the RWKV v7 language model, implemented with the Burn deep learning framework☆14Jun 9, 2025Updated last year
- Reinforcement Learning Toolkit for RWKV.(v6,v7,ARWKV) Distillation,SFT,RLHF(DPO,ORPO), infinite context training, Aligning. Exploring the…☆64Sep 19, 2025Updated 10 months ago
- Mini Model Daemon☆13Nov 9, 2024Updated last year
- RWKV centralised docs for the community☆35Jan 17, 2026Updated 6 months ago
- Language modeling with linear-cost context☆119Sep 25, 2025Updated 9 months ago
- Inference RWKV with multiple supported backends.☆94Jul 17, 2026Updated last week
- ☆21Jun 13, 2024Updated 2 years ago
- burn inference and training of dragon models 🔥🐉☆15Apr 12, 2026Updated 3 months ago
- Synthetic pretraining data by rephrasing the web☆24Jun 5, 2026Updated last month
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Overflow Prevention Enhances Long-Context Recurrent LLMs (COLM 2025)☆18Jul 8, 2025Updated last year
- An Ultra-Long Output Reinforcement Learning Approach☆23Jul 31, 2025Updated 11 months ago
- A Nanofactory Roadmap 2: Improved Proposal for a Comprehensive Diamondoid Nanofactory Development Program☆18Jul 24, 2025Updated last year
- ☆12Dec 21, 2024Updated last year
- Official Chinese documentation for RWKV | RWKV官方中文文档☆15Jun 10, 2026Updated last month
- A 20M RWKV v6 can do nonogram☆13Oct 18, 2024Updated last year
- Efficient PScan implementation in PyTorch☆17Jan 2, 2024Updated 2 years ago
- RWKV v5,v6 LoRA Trainer on Cuda and Rocm Platform. RWKV is a RNN with transformer-level LLM performance. It can be directly trained like …☆13Mar 24, 2024Updated 2 years ago
- An MLX implementation of Meta AI's ESM-2 protein language model☆16Aug 16, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆20Mar 11, 2025Updated last year
- A varitation graph tool☆10Dec 23, 2019Updated 6 years ago
- An official implementation of Random Policy Valuation is Enough for LLM Reasoning with Verifiable Rewards☆36Oct 3, 2025Updated 9 months ago
- An High-resolution implementation of HiFi-GAN Vocoder for Voice Conversion.☆32Apr 10, 2023Updated 3 years ago
- State tuning tunes the state☆35Feb 12, 2025Updated last year
- ☆19Sep 29, 2024Updated last year
- Diffusion Language Models For Code Infilling Beyond Fixed-size Canvas☆118Feb 3, 2026Updated 5 months ago