[ICML 26 Spotlight] Code for paper "VideoKR: Towards Knowledge- and Reasoning-Intensive Video Understanding"
☆21Jun 5, 2026Updated 4 months ago
Alternatives and similar repositories for VideoKR
Users that are interested in VideoKR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs☆20Feb 3, 2026Updated 8 months ago
- [EMNLP 2025] Code for paper "Table-R1: Inference-Time Scaling for Table Reasoning"☆34Jun 3, 2025Updated last year
- Data and Code for CVPR 2025 paper "MMVU: Measuring Expert-Level Multi-Discipline Video Understanding"☆77Feb 28, 2025Updated last year
- Data and code for ACL 2026 Paper "Rethinking Reasoning-Intensive Retrieval: Evaluating and Advancing Retrievers in Agentic Search Systems…☆22Apr 30, 2026Updated 5 months ago
- [Blog 1] Recording a bug of grpo_trainer in some R1 projects☆23Feb 23, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice☆89Feb 27, 2026Updated 7 months ago
- Code for "Skill-based Chain-of-Thoughts for Domain-Adaptive Video Reasoning [EMNLP 2025 Findings]"☆18Aug 27, 2025Updated last year
- ☆12Nov 19, 2022Updated 3 years ago
- PyTorch implementation of "Field Matching: an Electrostatic Paradigm to Generate and Transfer Data" (ICML 2025)☆21Jul 23, 2025Updated last year
- TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs☆138Updated this week
- 数据库实践课设:利用C#和SQL-Server实现简易的选课系统☆10Oct 11, 2020Updated 5 years ago
- [TKDE 2024] Robust Knowledge Adaptation for Dynamic Graph Neural Networks☆11Apr 11, 2024Updated 2 years ago
- This is the official implementation of paper "Multi-Prior Learning via Neural Architecture Search for Blind Face Restoration".☆13Jun 20, 2022Updated 4 years ago
- VideoMathQA is a benchmark designed to evaluate mathematical reasoning in real-world educational videos☆25Sep 5, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Analysis code for Neurips 2025 paper "SciArena: An Open Evaluation Platform for Foundation Models in Scientific Literature Tasks"☆57Aug 6, 2025Updated last year
- HoTPP: An Event Sequence Prediction Benchmark☆30Aug 27, 2026Updated last month
- Data and code for ACL 2023 paper "RobuT: A Systematic Study of Table QA Robustness Against Human-Annotated Adversarial Perturbations"☆15Feb 8, 2024Updated 2 years ago
- OpenClaw-RL: Personalize openclaw simply by talking to it☆16Feb 26, 2026Updated 7 months ago
- ☆29Aug 9, 2025Updated last year
- Official implementation of paper MiAD: Mirage Atom Diffusion for De Novo Crystal Generation☆33Jan 17, 2026Updated 8 months ago
- Repository for "Generative Flow Networks as Entropy-Regularized RL" (AISTATS-2024, Oral)☆41Apr 21, 2024Updated 2 years ago
- RWKV is an RNN with transformer-level LLM performance. It can be directly trained like a GPT (parallelizable). So it's combining the best…☆63Mar 17, 2025Updated last year
- ☆17May 19, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Bidirectional Likelihood Estimation with Multi-Modal Large Language Models for Text-Video Retrieval (ICCV 2025 Highlight)☆28Aug 1, 2025Updated last year
- Skoltech NLA 2024 course.☆38Dec 10, 2024Updated last year
- Official repository for the FIRM Reward series☆49Aug 25, 2026Updated last month
- This repo contains evaluation code for the paper "AV-Odyssey: Can Your Multimodal LLMs Really Understand Audio-Visual Information?"☆31Dec 23, 2024Updated last year
- ☆19Jun 29, 2025Updated last year
- UniAVGen: Unified Audio and Video Generation with Asymmetric Cross-Modal Interactions☆59Dec 16, 2025Updated 9 months ago
- Conference timeline browser built with Next.js☆16Apr 5, 2026Updated 6 months ago
- [NeurIPS 2025] VideoRFT: Incentivizing Video Reasoning Capability in MLLMs via Reinforced Fine-Tuning☆66Jan 6, 2026Updated 8 months ago
- docker-compose 一键式搭建 WordPress 个人博客☆21Nov 26, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official implementation of "NinA: Normalizing Flows in Action. Training VLA Models with Normalizing Flows"☆19Sep 22, 2025Updated last year
- ☆19Apr 21, 2026Updated 5 months ago
- ☆23Feb 26, 2024Updated 2 years ago
- Repository for the CVPR23 paper Re^2TAL☆13Nov 21, 2025Updated 10 months ago
- ☆55Dec 10, 2025Updated 9 months ago
- ☆157Apr 27, 2026Updated 5 months ago
- [ICLR 2026] "VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning?", Yuanxin Liu, Kun Ouyang, Haoning Wu, Yi Liu, L…☆41Jan 30, 2026Updated 8 months ago