WorldCache: Content-Aware Caching for Accelerated Video World Models
β27Jun 28, 2026Updated 3 months ago
Alternatives and similar repositories for WorldCache
Users that are interested in WorldCache are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code of the paper "VideoMolmo: Spatio-Temporal Grounding meets Pointing"β57Jul 5, 2025Updated last year
- [ACCV 2024] ObjectCompose: Evaluating Resilience of Vision-Based Models on Object-to-Background Compositional Changes πππβ37Jan 21, 2025Updated last year
- Language Grounded Single Source Domain Generalization in Medical Image Segmentation [ISBI2024]β34Oct 27, 2024Updated last year
- β41Jan 9, 2025Updated last year
- [ICLR 2026 π₯] Dr.LLM: Dynamic Layer Routing in LLMsβ58Apr 24, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [NAACL 2025 π₯] CAMEL-Bench is an Arabic benchmark for evaluating multimodal models across eight domains with 29,000 questions.β38Apr 17, 2025Updated last year
- ICLR 2026: Agent-X Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasksβ46Apr 28, 2026Updated 5 months ago
- A new multi-task learning framework using Vision Transformersβ11Jun 19, 2024Updated 2 years ago
- [CVPRW-25 MMFM] Official repository of paper titled "How Good is my Video LMM? Complex Video Reasoning and Robustness Evaluation Suite foβ¦β50Aug 23, 2024Updated 2 years ago
- β11Oct 29, 2024Updated last year
- AIN - The First Arabic Inclusive Large Multimodal Model. It is a versatile bilingual LMM excelling in visual and contextual understandingβ¦β55Mar 13, 2025Updated last year
- [BMVC 2024] On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Modelsβ15Nov 1, 2024Updated last year
- Learnable Weight Initialization for Volumetric Medical Image Segmentation [Elsevier AIM2024]β22Oct 27, 2024Updated last year
- VideoMathQA is a benchmark designed to evaluate mathematical reasoning in real-world educational videosβ25Sep 5, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [MICCAI 2024 π₯] HLSS, the first study to explore hierarchical information inherent in histopathology images and their language descriptiβ¦β28Aug 5, 2024Updated 2 years ago
- Official implementation of the paper "TOKENTRIM: INFERENCE-TIME TOKEN PRUNING FOR AUTOREGRESSIVE LONG VIDEO GENERATION"β15Feb 8, 2026Updated 7 months ago
- Self Evolving Large Multimodal Models with Continuous Rewardsβ27Sep 5, 2026Updated last month
- A Large Multimodal Model for Remote Sensing Change Description (IGARSS 2025)β22Dec 17, 2025Updated 9 months ago
- Official repository for "Boosting Adversarial Transferability using Dynamic Cues " (ICLR 2023)β20Aug 24, 2023Updated 3 years ago
- Unifying Visual Localization and Scene Recognition on Panoramic Annular Lensβ13May 18, 2020Updated 6 years ago
- repo for paper titled: Towards Realistic Zero-Shot Classification via Self Structural Semantic Alignment (AAAI'24 Oral)β25May 16, 2024Updated 2 years ago
- A command line interface app that allows EJUST students to manage all their stuff.β12Dec 4, 2022Updated 3 years ago
- Spatial Spectral Machine Learningβ14Oct 15, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR'26 Demo] Mobile-O: Unified Multimodal Understanding and Generation on Mobile Deviceβ159Apr 13, 2026Updated 5 months ago
- [NAACL'25] Contains code and documentation for our VANE-Bench paper.β24Aug 19, 2025Updated last year
- [ACL 2025 π₯] A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understandingβ79Aug 10, 2026Updated last month
- [NeurIPS 2025] Official PyTorch implementation of "Token Bottleneck: One Token to Remember Dynamics"β32Feb 2, 2026Updated 8 months ago
- [EMNLP'23] ClimateGPT: a specialized LLM for conversations related to Climate Change and Sustainability topics in both English and Arabiβ¦β80Sep 24, 2024Updated 2 years ago
- Official repository of paper titled "D3Former: Debiased Dual Distilled Transformer for Incremental Learning".β25Jul 10, 2023Updated 3 years ago
- [CVPR 2023] Bridging Precision and Confidence: A Train-Time Loss for Calibrating Object Detectionβ31Jun 21, 2023Updated 3 years ago
- We introduce new approach, Token Reduction using CLIP Metric (TRIM), aimed at improving the efficiency of MLLMs without sacrificing theirβ¦β23Jan 11, 2026Updated 8 months ago
- Code for Open3DTrack: Towards Open-Vocabulary 3D Multi-Object Trackingβ36Mar 14, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- DUET-VLM: Dual stage Unified Efficient Token reduction for VLM Training and Inferenceβ26May 21, 2026Updated 4 months ago
- [ICML 2025] Official PyTorch implementation of "NegMerge: Sign-Consensual Weight Merging for Machine Unlearning"β16Nov 25, 2025Updated 10 months ago
- Video-CoM: Interactive Video Reasoning via Chain of Manipulationsβ23Sep 5, 2026Updated last month
- Evaluation tool for the LILocBench benchmark challengeβ28Aug 8, 2025Updated last year
- [ICCV 2025] Scene Coordinate Reconstruction Priorsβ22Nov 10, 2025Updated 10 months ago
- Mobile-VideoGPT: Fast and Accurate Video Understanding Language Modelβ142Aug 6, 2025Updated last year
- β21Aug 2, 2026Updated 2 months ago