☆29May 13, 2025Updated last year
Alternatives and similar repositories for dymu
Users that are interested in dymu are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2025] DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models☆86Apr 16, 2026Updated 4 months ago
- Implementation of LaViC (KDD 2025)☆13Jun 1, 2025Updated last year
- The official source code for "Vision Language Model is NOT All You Need: Augmentation Strategies for Molecule Language Model".☆14Jul 23, 2024Updated 2 years ago
- iLLaVA: An Image is Worth Fewer Than 1/3 Input Tokens in Large Multimodal Models (ICLR2026)☆23Jun 24, 2026Updated last month
- ☆20Jul 24, 2026Updated 3 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- PyTorch code for the CVPR'23 paper: "ConStruct-VL: Data-Free Continual Structured VL Concepts Learning"☆14Feb 5, 2024Updated 2 years ago
- Open-source RL Framework with Online Teacher-Student Distillation☆22Mar 5, 2026Updated 5 months ago
- [EMNLP 2025 Oral] Official codebase for Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors.☆18Sep 7, 2025Updated 11 months ago
- ☆15Apr 25, 2025Updated last year
- Implementation for <Orthogonal Over-Parameterized Training> in CVPR'21.☆22Jul 16, 2021Updated 5 years ago
- [KDD 2023] Ball Trajectory Inference from Multi-Agent Sports Contexts Using Set Transformer and Hierarchical Bi-LSTM☆33Feb 3, 2026Updated 6 months ago
- [Technical Report] Official PyTorch implementation code for realizing the technical part of Phantom of Latent representing equipped with …☆63Oct 9, 2024Updated last year
- a training-free approach to accelerate ViTs and VLMs by pruning redundant tokens based on similarity☆44May 24, 2025Updated last year
- The official source code for [2023 NeurIPS] " Density of States Prediction of Crystalline Materials via Prompt-guided Multi-Modal Transfo…☆30Oct 15, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is an official implementation of our work, Select and Distill: Selective Dual-Teacher Knowledge Transfer for Continual Learning on V…☆17Sep 24, 2025Updated 10 months ago
- LLaVA-PruMerge: Adaptive Token Reduction for Efficient Large Multimodal Models☆174Mar 8, 2026Updated 5 months ago
- ☆14Sep 22, 2025Updated 10 months ago
- [ICML 2026] HiDe: Rethinking The Zoom-IN method in High Resolution MLLMs via Hierarchical Decoupling☆30May 2, 2026Updated 3 months ago
- A paper list of some recent works about Token Compress for Vit and VLM☆946Aug 10, 2026Updated last week
- The code for paper: "DC-Net: Divide-and-Conquer for Salient Object Detection"☆22Aug 30, 2024Updated last year
- Pytorch Implementation of the Model from "MIRASOL3B: A MULTIMODAL AUTOREGRESSIVE MODEL FOR TIME-ALIGNED AND CONTEXTUAL MODALITIES"☆26Jan 27, 2025Updated last year
- Official implementation of the paper: "A deeper look at depth pruning of LLMs"☆15Jul 24, 2024Updated 2 years ago
- Missing Modality Generation for Recommendaton☆38Aug 1, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ACM MM 2025] ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models☆18Jul 15, 2025Updated last year
- [EMNLP 2023] TESTA: Temporal-Spatial Token Aggregation for Long-form Video-Language Understanding☆50Jan 9, 2024Updated 2 years ago
- Research work aimed at addressing the problem of modeling infinite-length context☆52Dec 18, 2025Updated 8 months ago
- [ICML 2024] When Linear Attention Meets Autoregressive Decoding: Towards More Effective and Efficient Linearized Large Language Models☆35Jun 12, 2024Updated 2 years ago
- [NeurIPS 2025] Backpropagation-Free Test-Time Adaptation via Probabilistic Gaussian Alignment☆22Mar 18, 2026Updated 5 months ago
- Quantized Evolution Strategies: High-precision Fine-tuning of Quantized LLMs at Low-precision Cost☆20May 14, 2026Updated 3 months ago
- [EMNLP 2024] Official PyTorch implementation code for realizing the technical part of Traversal of Layers (TroL) presenting new propagati…☆99Jun 23, 2024Updated 2 years ago
- VisionSelector: End-to-End Learnable Visual Token Compression for Efficient Multimodal LLMs☆65Aug 9, 2026Updated last week
- The official code for the paper: LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs☆122Jul 1, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Official Repository for NeurIPS'25 Paper "Tool-Augmented Spatiotemporal Reasoning for Streamlining Video Question Answering Task"☆23May 18, 2026Updated 3 months ago
- Solving Token Gradient Conflict in Mixture-of-Experts for Large Vision-Language Model☆13Feb 11, 2025Updated last year
- ☆17Aug 1, 2025Updated last year
- [CVPR 2026] VideoSeek: Long-Horizon Video Agent with Tool-Guided Seeking☆67Mar 23, 2026Updated 4 months ago
- ☆15Apr 6, 2026Updated 4 months ago
- ☆29Jul 16, 2024Updated 2 years ago
- Code for MERL's ECCV 2022 paper on Cross-Modal Knowledge Transfer Without Task-Relevant Source Data☆11Jul 19, 2022Updated 4 years ago