Compress and Attend Transformers (CATs) 😸
☆23Aug 9, 2026Updated 3 weeks ago
Alternatives and similar repositories for cat-transformer
Users that are interested in cat-transformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- DeltaProduct is a new linear recurrent neural network architecture that uses products of generalized Householder matrices as state-transi…☆17Oct 13, 2025Updated 10 months ago
- ☆28Oct 2, 2025Updated 10 months ago
- Overcoming Long-Context Limitations of State-Space Models via Context-Dependent Sparse Attention (NeurIPS 2025)☆23Sep 30, 2025Updated 11 months ago
- HiCache: Hermite Polynomial-based Feature Cache for diffusion inference☆15Jul 29, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Flash-Linear-Attention models beyond language☆21Aug 28, 2025Updated last year
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- Official implementation: "AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation"☆20Oct 9, 2025Updated 10 months ago
- ☆17Jun 25, 2025Updated last year
- [CVPR2026 Findings] VHS: Verifier on Hidden States, an efficient inference-time scaling verification framework for DiT-based image genera…☆16Mar 25, 2026Updated 5 months ago
- mHC-lite: You Don’t Need 20 Sinkhorn-Knopp Iterations☆94Jan 12, 2026Updated 7 months ago
- ☆22Mar 3, 2026Updated 5 months ago
- [CVPR 2026 Findings] V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think☆57Apr 28, 2026Updated 4 months ago
- T5Voice is a lightweight PyTorch implementation of T5-based text-to-speech synthesis, supporting both streaming and non-streaming speech …☆28Nov 7, 2025Updated 9 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [ICLR 2026] FastVMT: This repo is the official implementation of "FastVMT: Eliminating Redundancy in Video Motion Transfer"☆26Feb 17, 2026Updated 6 months ago
- ☆24Nov 16, 2025Updated 9 months ago
- Official Implementation of MARS☆30Apr 21, 2026Updated 4 months ago
- ☆12Jul 25, 2023Updated 3 years ago
- ☆25Jun 19, 2025Updated last year
- ☆37Dec 31, 2025Updated 7 months ago
- ☆29Mar 10, 2026Updated 5 months ago
- CLIP-Art: Contrastive Pre-training for Fine-Grained Art Classification - 4th Workshop on Computer Vision for Fashion, Art, and Design☆28May 2, 2022Updated 4 years ago
- ☆26Aug 24, 2026Updated last week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Descript Audio Codec - VAE Variant (.dac-vae): High-Fidelity Audio Compression with Variational Autoencoder☆40Aug 30, 2025Updated last year
- Xmixers: A collection of SOTA efficient token/channel mixers☆29Sep 4, 2025Updated 11 months ago
- ☆48Dec 13, 2025Updated 8 months ago
- Try to replicate the architecture of MiniMaxTTS mentioned in it's technical report☆47Sep 2, 2025Updated 11 months ago
- Official repository for “Duo-Tok: Dual-Track Semantic Music Tokenizer for Vocal–Accompaniment Generation.”☆32Nov 26, 2025Updated 9 months ago
- Open implementation of Attention Residuals (Kimi Team, arXiv:2603.15031)☆83Apr 30, 2026Updated 4 months ago
- A large database of artificial neural network statistics during training☆15Dec 8, 2020Updated 5 years ago
- Inference for the STFT-VAE continuous audio codec (24kHz, 3.125Hz latent)☆47Jul 12, 2026Updated last month
- Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence☆69Nov 11, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆65Jun 12, 2025Updated last year
- Code for reproducing the results from "CrAM: A Compression-Aware Minimizer" accepted at ICLR 2023☆10Mar 1, 2023Updated 3 years ago
- Official code for paper:"Speaking Clearly: A Simplified Whisper-Based Codec for Low-Bitrate Speech Coding"☆38Jan 28, 2026Updated 7 months ago
- Official repository for ICML 2024 paper "MoRe Fine-Tuning with 10x Fewer Parameters"☆22Oct 14, 2025Updated 10 months ago
- Official repository for Parallax (Parameterized Local Linear Attention)☆68Jul 30, 2026Updated last month
- A series of models applying memory augmented neural networks to machine translation☆15May 3, 2018Updated 8 years ago
- An hardware-aware Efficient Implementation for "Mixture-of-Depths Attention".☆274May 6, 2026Updated 3 months ago