[ICML26] Official Repo for WorldCache: Accelerating World Models for Free via Heterogeneous Token Caching
☆43Jul 23, 2026Updated last month
Alternatives and similar repositories for WorldCache
Users that are interested in WorldCache are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026] EarlyTom: Early Token Compression Completes Fast Video Understanding☆37Jun 22, 2026Updated 2 months ago
- Implementation of RankE: End-to-End Discrete Text-to-Image Post-Training via Rank-Consistent Alignment☆22May 27, 2026Updated 3 months ago
- [arxiv 2025] SparseSSM: Efficient Selective Structured State Space Models Can Be Pruned in One-Shot☆23Oct 8, 2025Updated 11 months ago
- Webpage: https://chuny9743.github.io/AI4WaterEnv_Webpage/☆36Aug 22, 2026Updated 3 weeks ago
- [ICLR 2026] MergeMix: A Unified Augmentation Paradigm for Visual and Multi-Modal Understanding☆23Feb 27, 2026Updated 6 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [NeurIPS 2025] HoliTom: Holistic Token Merging for Fast Video Large Language Models☆85Oct 10, 2025Updated 11 months ago
- ☆43Jun 8, 2026Updated 3 months ago
- [CPAL 2026 oral] Offical implementation of "ROSE: Reordered SparseGPT for More Accurate One-Shot Large Language Models Pruning”☆17Jul 31, 2026Updated last month
- [ICML 2026] PASA: A Principled Embedding-Space Watermarking Approach for LLM-Generated Text under Semantic-Invariant Attacks☆24May 13, 2026Updated 4 months ago
- LVOmniBench: Pioneering Long Audio-Video Understanding Evaluation for Omnimodal LLMs☆42Apr 2, 2026Updated 5 months ago
- A list of awesome papers on compression and acceleration of Large Language Models (LLMs) and Multimodal Large Language Models (MLLMs).☆18May 12, 2026Updated 4 months ago
- CGGS: Consistency-Augmented Geometric Gaussian Splatting for Ego-Centric 3D Scene Generation (TIP 2026)☆26Jul 23, 2026Updated last month
- [arXiv 2026] dVoting: Fast Voting for dLLMs☆30Feb 13, 2026Updated 7 months ago
- [CVPR 2026] OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models☆108Apr 20, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICML 2026] Official Repo for Fast-SAM3D: 3Dfy Anything in Images but Faster☆195Jul 25, 2026Updated last month
- [ICLR 2026] Code for QuantVGGT: Quantized Visual Geometry Grounded Transformer☆121Mar 20, 2026Updated 5 months ago
- [CVPR 2026] ReasonMap: Towards Fine-Grained Visual Reasoning from Transit Maps☆87Jul 26, 2026Updated last month
- [ICCV 2025] QuantCache:Adaptive Importance-Guided Quantization with Hierarchical Latent and Layer Caching for Video Generation☆18Sep 26, 2025Updated 11 months ago
- [CVPR 2026 Oral] SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware Caching☆27Jun 5, 2026Updated 3 months ago
- [ICLR 2026] RewardMap: Tackling Sparse Rewards in Fine-grained Visual Reasoning via Multi-Stage Reinforcement Learning☆47Feb 22, 2026Updated 6 months ago
- A list of papers, docs, codes about diffusion distillation.This repo collects various distillation methods for the Diffusion model. Welc…☆41Dec 10, 2023Updated 2 years ago
- A list of papers, docs, codes about diffusion quantization.This repo collects various quantization methods for the Diffusion Models. Welc…☆25Feb 2, 2026Updated 7 months ago
- [TMLR 2026] Is Oracle Pruning the True Oracle?☆35Jul 1, 2026Updated 2 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆17Oct 5, 2025Updated 11 months ago
- [ICLR 2026] Official implementation of "Enhancing Multi-Image Understanding Through Delimiter Token Scaling"☆17Jul 10, 2026Updated 2 months ago
- (ICML-2025) Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers☆21Aug 13, 2025Updated last year
- [ICLR 2026 Oral] RAIN-Merging☆15Mar 9, 2026Updated 6 months ago
- ☆37May 21, 2026Updated 3 months ago
- Official implementation for WorldScore: A Unified Evaluation Benchmark for World Generation☆310Jul 23, 2026Updated last month
- [ICML 2025] This is the official PyTorch implementation of "🎵 HarmoniCa: Harmonizing Training and Inference for Better Feature Caching i…☆46Jul 10, 2025Updated last year
- [NeurIPS'25] dKV-Cache: The Cache for Diffusion Language Models☆136May 22, 2025Updated last year
- [ICML 2026] Stable Asynchrony: Variance-Controlled Off-Policy RL for LLMs☆34Apr 27, 2026Updated 4 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [Arxiv 2025] In-Video Instructions: Visual Signals as Generative Control☆46Nov 25, 2025Updated 9 months ago
- [NeurIPS'25] FreqExit: Enabling Early-Exit Inference for Visual Autoregressive Models via Frequency-Aware Guidance☆21Dec 15, 2025Updated 9 months ago
- ☆15Apr 3, 2026Updated 5 months ago
- Collection of recent works on AI Agents.☆17Jun 5, 2025Updated last year
- A Library for intra-GPU/Inter-SM parallelsim☆12Aug 7, 2026Updated last month
- Official implementation of "Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion"☆108Aug 25, 2026Updated 3 weeks ago
- Minute-long video generation at 24FPS.☆70Mar 28, 2026Updated 5 months ago