possibly useful materials for learning RWKV language model.
☆27Jun 8, 2023Updated 3 years ago
Alternatives and similar repositories for RWKV-howto
Users that are interested in RWKV-howto are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 一个简单的,由ChatGPT主导编写的api,使用简单的请求访问ChatRWKV☆15May 19, 2023Updated 3 years ago
- Code for the paper: "T-shape data and probabilistic remaining useful life prediction for Li-ion batteries using multiple non-crossing qua…☆10Aug 4, 2023Updated 3 years ago
- 一个用Apple Metal实现的Llama和通义千问大模型本地推理☆10Apr 26, 2024Updated 2 years ago
- All-in-one benchmarking platform for evaluating LLM.☆15Nov 12, 2025Updated 10 months ago
- Contains the code for the paper "Multi-Horizon Short-Term Load Forecasting Using Hybrid of LSTM and Modified Split Convolution"☆11Oct 28, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- BlinkDL's RWKV-v4 running in the browser☆48Mar 2, 2023Updated 3 years ago
- ☆17Aug 1, 2023Updated 3 years ago
- ☆19Dec 12, 2023Updated 2 years ago
- Ἀνατομή is a PyTorch library to analyze representation of neural networks☆13Jan 31, 2024Updated 2 years ago
- codes and plots for "Active-Dormant Attention Heads: Mechanistically Demystifying Extreme-Token Phenomena in LLMs"☆11Dec 30, 2024Updated last year
- Python implementation of AWarp algorithm☆14Aug 6, 2021Updated 5 years ago
- ☆32Mar 30, 2023Updated 3 years ago
- Fine-tuning RWKV-World model☆26Jun 6, 2023Updated 3 years ago
- 🎓Automatically Update CV Papers Daily using Github Actions (Update Every 12th hours)☆12May 17, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code for paper "DB-LSTM: Densely-Connected Bi-directional LSTM for Human Action Recognition"☆13Jul 1, 2022Updated 4 years ago
- ☆15Sep 1, 2025Updated last year
- Code for experiments on transformers using Markovian data.☆22Nov 22, 2024Updated last year
- Find context neurons in Pythia models.☆13Jun 13, 2023Updated 3 years ago
- A converter and basic tester for rwkv onnx☆44Jan 29, 2024Updated 2 years ago
- MLOps Model Factory is an end to end workflow that supports generating multiple models and used for deployment to any target.☆10May 9, 2024Updated 2 years ago
- LongAttn :Selecting Long-context Training Data via Token-level Attention☆15Jul 16, 2025Updated last year
- Least Squares Regression for subspace clustering☆11May 27, 2018Updated 8 years ago
- 一个基于Flask实现的RWKV_Role_Playing项目的API。☆32Jun 26, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The official implementation of HybridNorm: Towards Stable and Efficient Transformer Training via Hybrid Normalization☆19Mar 7, 2025Updated last year
- Implementation of Unified Embedding: Battle-Tested Feature Representations for Web-Scale ML Systems☆15Nov 11, 2023Updated 2 years ago
- Implementation of the dilated self attention as described in "LongNet: Scaling Transformers to 1,000,000,000 Tokens"☆13Jul 23, 2023Updated 3 years ago
- The Hessian of tall-skinny networks is easy to invert☆17Sep 5, 2026Updated 2 weeks ago
- Personal solutions to the Triton Puzzles☆22Jul 18, 2024Updated 2 years ago
- ☆43Mar 29, 2023Updated 3 years ago
- RWKV-v2-RNN trained on the Pile. See https://github.com/BlinkDL/RWKV-LM for details.☆67Sep 14, 2022Updated 4 years ago
- ☆24Oct 10, 2025Updated 11 months ago
- ☆12Sep 7, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆17Sep 27, 2022Updated 3 years ago
- RWKV (Receptance Weighted Key Value) is a RNN with Transformer-level performance☆40Feb 12, 2023Updated 3 years ago
- The Newton-Muon optimizer☆32Jun 5, 2026Updated 3 months ago
- ☆16Feb 6, 2024Updated 2 years ago
- PyTorch implementation of "Towards k-means-friendly spaces: Simultaneous deep learning and clustering," Bo Yang et al., 2017.☆17Jan 15, 2021Updated 5 years ago
- continous batching and parallel acceleration for RWKV6☆23Jun 28, 2024Updated 2 years ago
- Implementation and evaluation of Scaling Embedding Layers in Language Models research paper☆17Feb 2, 2026Updated 7 months ago