Compress and Attend Transformers (CATs) 😸
☆23Jul 13, 2026Updated last week
Alternatives and similar repositories for cat-transformer
Users that are interested in cat-transformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- DeltaProduct is a new linear recurrent neural network architecture that uses products of generalized Householder matrices as state-transi…☆15Oct 13, 2025Updated 9 months ago
- ☆28Oct 2, 2025Updated 9 months ago
- Overcoming Long-Context Limitations of State-Space Models via Context-Dependent Sparse Attention (NeurIPS 2025)☆23Sep 30, 2025Updated 9 months ago
- HiCache: Hermite Polynomial-based Feature Cache for diffusion inference☆15Jan 27, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [CVPR 2022] Official code for the paper: "A Stitch in Time Saves Nine: A Train-Time Regularizing Loss for Improved Neural Network Calibra…☆33Nov 9, 2022Updated 3 years ago
- Flash-Linear-Attention models beyond language☆21Aug 28, 2025Updated 10 months ago
- Official implementation: "AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation"☆19Oct 9, 2025Updated 9 months ago
- [CVPR2026 Findings] VHS: Verifier on Hidden States, an efficient inference-time scaling verification framework for DiT-based image genera…☆16Mar 25, 2026Updated 3 months ago
- mHC-lite: You Don’t Need 20 Sinkhorn-Knopp Iterations☆91Jan 12, 2026Updated 6 months ago
- ☆21Mar 3, 2026Updated 4 months ago
- [CVPR 2026 Findings] V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think☆56Apr 28, 2026Updated 2 months ago
- [ICLR 2026] FastVMT: This repo is the official implementation of "FastVMT: Eliminating Redundancy in Video Motion Transfer"☆26Feb 17, 2026Updated 5 months ago
- T5Voice is a lightweight PyTorch implementation of T5-based text-to-speech synthesis, supporting both streaming and non-streaming speech …☆28Nov 7, 2025Updated 8 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆24Nov 16, 2025Updated 8 months ago
- Official Implementation of MARS☆30Apr 21, 2026Updated 2 months ago
- ☆25Jun 19, 2025Updated last year
- ☆34Dec 31, 2025Updated 6 months ago
- ☆26Apr 30, 2026Updated 2 months ago
- Descript Audio Codec - VAE Variant (.dac-vae): High-Fidelity Audio Compression with Variational Autoencoder☆38Aug 30, 2025Updated 10 months ago
- ☆48Dec 13, 2025Updated 7 months ago
- Try to replicate the architecture of MiniMaxTTS mentioned in it's technical report☆47Sep 2, 2025Updated 10 months ago
- Official repository for “Duo-Tok: Dual-Track Semantic Music Tokenizer for Vocal–Accompaniment Generation.”☆32Nov 26, 2025Updated 7 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Open implementation of Attention Residuals (Kimi Team, arXiv:2603.15031)☆74Apr 30, 2026Updated 2 months ago
- Conditional Random Fields implemented as Lasagne layer☆10Jul 22, 2016Updated 9 years ago
- ☆24Mar 18, 2025Updated last year
- ☆13Aug 4, 2022Updated 3 years ago
- Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence☆68Nov 11, 2025Updated 8 months ago
- Inference for the STFT-VAE continuous audio codec (24kHz, 3.125Hz latent)☆43Jul 12, 2026Updated last week
- A collection of my personal Linux dot files.☆12Jul 28, 2020Updated 5 years ago
- ☆65Jun 12, 2025Updated last year
- Code for reproducing the results from "CrAM: A Compression-Aware Minimizer" accepted at ICLR 2023☆10Mar 1, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official code for paper:"Speaking Clearly: A Simplified Whisper-Based Codec for Low-Bitrate Speech Coding"☆37Jan 28, 2026Updated 5 months ago
- ☆57Oct 23, 2023Updated 2 years ago
- Official repository for Parallax (Parameterized Local Linear Attention)☆65Jul 7, 2026Updated 2 weeks ago
- Official repository for ICML 2024 paper "MoRe Fine-Tuning with 10x Fewer Parameters"☆22Oct 14, 2025Updated 9 months ago
- A series of models applying memory augmented neural networks to machine translation☆15May 3, 2018Updated 8 years ago
- An hardware-aware Efficient Implementation for "Mixture-of-Depths Attention".☆274May 6, 2026Updated 2 months ago
- Tidy Tunes is an easy-to-use pipeline for mining high-quality audio data for speech generation models. To do so, it chains multiple open …☆23May 19, 2026Updated 2 months ago