Official repository of paper "LOVE-R1: Advancing Long Video Understanding with Adaptive Zoom-in Mechanism via Multi-Step Reasoning"
☆24Nov 1, 2025Updated 9 months ago
Alternatives and similar repositories for LOVE-R1
Users that are interested in LOVE-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- (ICCV2025) Official repository of paper "ViSpeak: Visual Instruction Feedback in Streaming Videos"☆54Jul 1, 2025Updated last year
- (ECCV2026) Official repository of paper "IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Gene…☆30Jul 1, 2026Updated last month
- (ICML 2026) Seg-ReSearch: Segmentation with Interleaved Reasoning and External Search☆50Updated this week
- [CVPR2026] Official repository of paper "CycleManip: Enabling Cyclic Task Manipulation via Effective Historical Perception and Understand…☆25Feb 21, 2026Updated 5 months ago
- Collect papers about Mamba (a selective state space model).☆15Aug 6, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CoRL2025] Official repository of paper "TypeTele: Releasing Dexterity in Teleoperation by Dexterous Manipulation Types".☆28Dec 3, 2025Updated 8 months ago
- ☆81Nov 24, 2024Updated last year
- The official implementation of our work Hawkeye: Discovering and Grounding Implicit Anomalous Sentiment in Recon-videos via Scene-enhanc…☆13Oct 14, 2024Updated last year
- [CVPR2025] Hybrid-Level Instruction Injection for Video Token Compression in Multi-modal Large Language Models☆21Apr 30, 2025Updated last year
- [ICLR 2026] ReWatch-R1: Boosting Complex Video Reasoning in Large Vision-Language Models through Agentic Data Synthesis☆30Mar 27, 2026Updated 4 months ago
- [ICLR 2026] Official repo for "FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting"☆50Oct 9, 2025Updated 10 months ago
- Enhancing Representations through Heterogeneous Self-Supervised Learning (TPAMI 2025)☆15May 2, 2025Updated last year
- (ECCV 2024) Official repository of paper "Bridge Past and Future: Overcoming Information Asymmetry in Incremental Object Detection"☆20Mar 26, 2025Updated last year
- [NeurIPS 2025] Panoptic Captioning: An Equivalence Bridge for Image and Text☆38Jan 31, 2026Updated 6 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- (ECCV 2024) Official repository of paper "EgoExo-Fitness: Towards Egocentric and Exocentric Full-Body Action Understanding"☆38Apr 8, 2025Updated last year
- (NeurIPS 2024) Official repository of paper "Frozen-DETR: Enhancing DETR with Image Understanding from Frozen Foundation Models"☆34Mar 22, 2025Updated last year
- Joint Selection for Large-Scale Pre-Training Data via Policy Gradient-based Mask Learning☆21Jan 4, 2026Updated 7 months ago
- The official code of "Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning"☆102Oct 15, 2025Updated 9 months ago
- (CVPR 2026) Long-RVOS: A Comprehensive Benchmark for Long-term Referring Video Object Segmentation☆38Feb 28, 2026Updated 5 months ago
- ☆19Jun 18, 2024Updated 2 years ago
- Offical implementation of "Re-Aligning Language to Visual Objects with an Agentic Workflow"☆34Apr 20, 2025Updated last year
- (ICML 2026) Official repository of paper "ObjEmbed: Towards Universal Multimodal Object Embeddings"☆55May 18, 2026Updated 2 months ago
- ☆40Nov 5, 2025Updated 9 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [NIPS2025] VideoChat-R1 & R1.5: Enhancing Spatio-Temporal Perception and Reasoning via Reinforcement Fine-Tuning☆269Oct 18, 2025Updated 9 months ago
- The official code for the paper: LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs☆122Jul 1, 2025Updated last year
- ☆163Jul 31, 2025Updated last year
- [ECCV 26] Video Streaming Thinking☆122Jul 28, 2026Updated 2 weeks ago
- The code implementation for the paper "DreamLifting: A Plug-in Module Lifting MV Diffusion Models for 3D Asset Generation".☆30Sep 1, 2025Updated 11 months ago
- [ECCV 2026] Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?☆97Jul 13, 2025Updated last year
- [ACM MM 2025] ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models☆18Jul 15, 2025Updated last year
- [AAAI2025] This is the official PyTorch codes for the paper: "DiffRetouch: Using Diffusion to Retouch on the Shoulder of Experts"☆25Jun 16, 2025Updated last year
- [NeurIPS 2025] Official repository of the paper "Unlocking Aha Moments via Reinforcement Learning: Advancing Collaborative Visual Compreh…☆23Sep 27, 2025Updated 10 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆154Nov 17, 2025Updated 8 months ago
- Video-R1: Reinforcing Video Reasoning in MLLMs [🔥the first paper to explore R1 for video]☆886Dec 14, 2025Updated 7 months ago
- [ECCV 2026] VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs☆17Feb 3, 2026Updated 6 months ago
- [ICML 2025 Oral] This is the official repository of the paper "What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensi…☆23Jun 12, 2025Updated last year
- ViCaS: A Dataset for Combining Holistic and Pixel-level Video Understanding using Captions with Grounded Segmentation (CVPR'25)☆21Apr 2, 2025Updated last year
- (ICCV 2023) Official implementation of Rectified Straight Through Estimator (ReSTE).☆34Sep 20, 2024Updated last year
- [ICCV 2025] Official Implementation of RefEdit: A Benchmark and Method for Improving Instruction-based Image Editing Model for Referring …☆20Jun 27, 2025Updated last year