VidKV: Plug-and-Play 1.x-Bit KV Cache Quantization for Video Large Language Models
☆25Mar 26, 2025Updated last year
Alternatives and similar repositories for VidKV
Users that are interested in VidKV are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2025] DyCoke: Dynamic Compression of Tokens for Fast Video Large Language Models☆114Nov 22, 2025Updated 8 months ago
- Self-training LLaVA for medical☆16Nov 3, 2024Updated last year
- [arxiv 2025] SparseSSM: Efficient Selective Structured State Space Models Can Be Pruned in One-Shot☆22Oct 8, 2025Updated 10 months ago
- ☆121Jun 2, 2026Updated 2 months ago
- ☆23Sep 3, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆65Jun 16, 2025Updated last year
- [ICDM 2023] Momentum is All You Need for Data-Driven Adaptive Optimization☆26Mar 30, 2024Updated 2 years ago
- ThinK: Thinner Key Cache by Query-Driven Pruning☆30Jun 2, 2026Updated 2 months ago
- ☆38Jun 2, 2026Updated 2 months ago
- [NeurIPS-2021] Slow Learning and Fast Inference: Efficient Graph Similarity Computation via Knowledge Distillation☆43Mar 24, 2023Updated 3 years ago
- [NeurIPS 2025] HoliTom: Holistic Token Merging for Fast Video Large Language Models☆84Oct 10, 2025Updated 10 months ago
- [ICLR'21] Neural Pruning via Growing Regularization (PyTorch)☆82Jul 15, 2021Updated 5 years ago
- [Preprint] Why is the State of Neural Network Pruning so Confusing? On the Fairness, Comparison Setup, and Trainability in Network Prunin…☆41Sep 9, 2025Updated 11 months ago
- [NeurIPS'22] What Makes a "Good" Data Augmentation in Knowledge Distillation -- A Statistical Perspective☆37Dec 15, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ACL 2025] PruneVid: Visual Token Pruning for Efficient Video Large Language Models☆71May 15, 2025Updated last year
- Source code and notebooks to reproduce experiments and benchmarks on Bias Faces in the Wild (BFW).☆52Jun 9, 2026Updated 2 months ago
- ☆13Mar 28, 2025Updated last year
- KV cache compression via sparse coding☆18Oct 26, 2025Updated 9 months ago
- ☆12May 15, 2025Updated last year
- ☆13Aug 17, 2020Updated 6 years ago
- [ICLR'23] Trainability Preserving Neural Pruning (PyTorch)☆34May 21, 2023Updated 3 years ago
- Explainable Face Recognition ECCV 2020 Paper code and dataset repository☆63Apr 28, 2021Updated 5 years ago
- [TCSVT 2025] Niagara: Normal-Integrated Geometric Affine Field for Scene Reconstruction from a Single View☆100Dec 15, 2025Updated 8 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Code for Neural Relational Inference with Efficient Message Passing Mechanisms (AAAI 2021).☆21May 9, 2021Updated 5 years ago
- AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference☆21Jan 24, 2025Updated last year
- CastleHill: Separable Causal Diffusion / Varitaion Flow Maps for LTX-2 long-form video generation☆15May 19, 2026Updated 2 months ago
- PrefixKV: Adaptive Prefix KV Cache is What Vision Instruction-Following Models Need for Efficient Generation [NeurIPS 2025]☆19Oct 11, 2025Updated 10 months ago
- Official Implementation (Pytorch) of the "Generative Subgraph Retrieval for Knowledge Graph-Grounded Dialog Generation", EMNLP 2024 (main…☆12Mar 10, 2025Updated last year
- [TKDE 2024, CIKM 2022] SLA²P: Self-supervised Anomaly Detection with Adversarial Perturbation.☆39Dec 26, 2024Updated last year
- The official code of "Mano: Restriking Manifold Optimization for LLM Training".☆25Jun 1, 2026Updated 2 months ago
- Official Implementation of "Gumiho: A Hybrid Architecture to Prioritize Early Tokens in Speculative Decoding" (ICML'25)☆33May 14, 2026Updated 3 months ago
- Page for the CVPR 2023 Tutorial - Efficient Neural Networks: From Algorithm Design to Practical Mobile Deployments☆12Jun 30, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Tuning-Free Image Editing with Fidelity and Editability via Unified Latent Diffusion Model☆13Dec 29, 2024Updated last year
- Unified Controllable Visual Generation Model☆662Jun 2, 2026Updated 2 months ago
- LTX-Video-Trainer-GUI 是为LTX视频lora模型训练提供的GUI工具,支持通过简单的界面训练 LoRA 模型用于视频生成。本训练器提供了直观的 GUI 界面,使用户能够轻松设置和启动训练流程,无需编写复杂代码。☆13Jul 18, 2025Updated last year
- [ICDM 2022] Making Reconstruction-based Method Great Again for Video Anomaly Detection (PyTorch)☆40Mar 25, 2024Updated 2 years ago
- 复旦研究生入学教育测试☆22Aug 28, 2025Updated 11 months ago
- HyperCUT: Video Sequence from a Single Blurry Image using Unsupervised Ordering (CVPR'23)☆14Nov 4, 2025Updated 9 months ago
- Official implementation of Sync-LoRA☆27Jun 25, 2026Updated last month