[ICLR 2026 π₯] Official pytorch implementation for "Attention Is All You Need for KV Cache in Diffusion LLMs"
β42Jul 13, 2026Updated last month
Alternatives and similar repositories for Elastic-Cache
Users that are interested in Elastic-Cache are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026 π₯] Time Blindness: Why Video-Language Models Can't See What Humans Can?β67Jan 28, 2026Updated 6 months ago
- Official code of the paper "VideoMolmo: Spatio-Temporal Grounding meets Pointing"β57Jul 5, 2025Updated last year
- [ICML 2026] Official implementation of "FOCUS: DLLMs Know How to Tame Their Compute Bound".β18Aug 2, 2026Updated 2 weeks ago
- [ACL 2026 π₯] CASS: Nvidia to AMD Transpilation with Data, Models, and Benchmarkβ37Updated this week
- [ACL 2025 π₯] A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understandingβ79Aug 10, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Sequential Diffusion Language Model (SDLM) enhances pre-trained autoregressive language models by adaptively determining generation lengtβ¦β98Dec 27, 2025Updated 7 months ago
- β18Aug 11, 2020Updated 6 years ago
- [ICLR'26] Official code of paper "d2Cache: Accelerating Diffusion-based LLMs via Dual Adaptive Caching"β159May 14, 2026Updated 3 months ago
- β29Oct 16, 2025Updated 10 months ago
- Official implementation of "Diffusion Language Models Know the Answer Before Decoding"β61Apr 28, 2026Updated 3 months ago
- [NeurIPS'25] dKV-Cache: The Cache for Diffusion Language Modelsβ135May 22, 2025Updated last year
- CodeRosetta: Pushing the Boundaries of Unsupervised Code Translation for Parallel Programmingβ11Nov 18, 2024Updated last year
- WorldCache: Content-Aware Caching for Accelerated Video World Modelsβ23Jun 28, 2026Updated last month
- The official GitHub repo for the survey paper "A Survey on Diffusion Language Models".β1,177Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Flexible and Pluggable Serving Engine for Diffusion LLMsβ152Jul 13, 2026Updated last month
- β38Jul 21, 2025Updated last year
- [ICLR 2026 π₯] Dr.LLM: Dynamic Layer Routing in LLMsβ57Apr 24, 2026Updated 3 months ago
- Official Implementation of MARSβ30Apr 21, 2026Updated 3 months ago
- Language Models for Code Completion: a Practical Evaluationβ13Jan 19, 2024Updated 2 years ago
- An official repository for GPTailorβ19Jun 29, 2025Updated last year
- [ICLR 2026] AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Sizeβ16Jan 28, 2026Updated 6 months ago
- OpenDLM is an open-source library focused on sampling algorithms for Diffusion Language Models (DLMs).β15Aug 5, 2025Updated last year
- [ICLR 2025] π CodeMMLU Evaluator: A framework for evaluating LM models on CodeMMLU MCQs benchmark.β29Apr 21, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- β15Jun 5, 2025Updated last year
- Voronoi-Based Foveated Volume Renderingβ10Sep 30, 2021Updated 4 years ago
- The official implementation of MaskGRPO: Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models. (ICLR 2026, arxivβ¦β19Jan 27, 2026Updated 6 months ago
- β10Sep 13, 2022Updated 3 years ago
- this repo attemps to reproduce DSOD: Learning Deeply Supervised Object Detectors from Scratch use gluon reimplementationβ14Aug 18, 2018Updated 8 years ago
- DMax: Aggressive Parallel Decoding for dLLMsβ127Jul 5, 2026Updated last month
- Repository for MetaVC -- A Meta Local Search Framework For Minimum Vertex Cover (MinVC)β10Jan 15, 2022Updated 4 years ago
- Offical implementation of the paper "Rhizomorph: The Coordinated Function of Shoots and Roots"β15Nov 21, 2023Updated 2 years ago
- β37Feb 4, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- β10Feb 28, 2023Updated 3 years ago
- β16May 23, 2024Updated 2 years ago
- LLM Evaluation Framework for Hardware Design Using Python-Embedded DSLsβ18Aug 26, 2024Updated last year
- [NeurIPS 2025] Scaling Speculative Decoding with Lookahead Reasoningβ69Oct 31, 2025Updated 9 months ago
- [CVPR 2022] DiSparse: Disentangled Sparsification for Multitask Model Compressionβ13Sep 6, 2022Updated 3 years ago
- Federated Conformal Prediction with Quantile-of-Quantiles (FedCP-QQ)β11May 6, 2026Updated 3 months ago
- Are gradient information useful for pruning of LLMs?β48Aug 23, 2025Updated 11 months ago