[EMNLP 2024] A Video Chat Agent with Temporal Prior
☆34Mar 2, 2025Updated last year
Alternatives and similar repositories for VideoTGB
Users that are interested in VideoTGB are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Repo for ReflectEvo☆22Jun 16, 2025Updated last year
- ☆14Dec 19, 2024Updated last year
- ☆14Feb 26, 2024Updated 2 years ago
- Official Repo of LangSuitE☆87Aug 15, 2024Updated 2 years ago
- ☆12Mar 4, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Weakly Supervised Gaussian Contrastive Grounding with Large Multimodal Models for Video Question Answering [ACM MM'24]☆10Jul 22, 2024Updated 2 years ago
- [ICML 2025] |TokenSwift: Lossless Acceleration of Ultra Long Sequence Generation☆126May 19, 2025Updated last year
- Contrastive Video Question Answering via Video Graph Transformer (IEEE T-PAMI'23)☆20Mar 9, 2024Updated 2 years ago
- The official source code for "Boosting LLM Agents with Recursive Contemplation for Effective Deception Handling" (ACL 2024, Findings)☆15Aug 12, 2024Updated 2 years ago
- ☆14Dec 16, 2023Updated 2 years ago
- ☆16Apr 12, 2024Updated 2 years ago
- [CVPR 2025] OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts☆20Apr 2, 2025Updated last year
- MR. Video: MapReduce is the Principle for Long Video Understanding☆32Jun 18, 2026Updated 3 months ago
- ☆32Feb 23, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆12Dec 15, 2023Updated 2 years ago
- [CVPR 2025] OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts☆28Jul 14, 2026Updated 2 months ago
- Large Language Models are Temporal and Causal Reasoners for Video Question Answering (EMNLP 2023)☆77Mar 26, 2025Updated last year
- EMNLP 2026 | Official Repository of UltraVoice☆67Aug 30, 2026Updated 3 weeks ago
- IMG: Calibrating Diffusion Models via Implicit Multimodal Guidance, ICCV 2025☆30Oct 1, 2025Updated 11 months ago
- ☆21Aug 23, 2026Updated last month
- [ECCV 2022] Learning to Weight Samples for Dynamic Early-exiting Networks☆38Sep 28, 2023Updated 3 years ago
- ☆33Nov 14, 2025Updated 10 months ago
- ☆11May 2, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆153Apr 16, 2025Updated last year
- Repository of "Train Once, Get a Family: State-Adaptive Balances for Offline-to-Online Reinforcement Learning" (NeurIPS 2023 Spotlight)☆41Oct 30, 2023Updated 2 years ago
- 中国历年GDP和人口数据可视化☆13Jan 18, 2023Updated 3 years ago
- [CVPR 2024] Adapting Short-Term Transformers for Action Detection in Untrimmed Videos☆11Jun 11, 2024Updated 2 years ago
- Official implementation of Dynamic Perceiver☆44Nov 16, 2023Updated 2 years ago
- [WACV 2025] Exploiting VLM Localizability and Semantics for Open Vocabulary Action Detection☆17Mar 23, 2025Updated last year
- A Data Visualization project on the French traffic accidents database☆19Aug 27, 2019Updated 7 years ago
- IMAGEimate is an end-to-end pipeline to create realistic animatable 3D avatars from a single image using neural networks☆13Dec 9, 2021Updated 4 years ago
- (2024CVPR) MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding☆354Jul 19, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Repository to generate CLEVR-Dialog: A diagnostic dataset for Visual Dialog☆50Feb 18, 2020Updated 6 years ago
- A simplified version of MPN☆13May 21, 2021Updated 5 years ago
- ☆39Jul 8, 2025Updated last year
- 🔥 Regeneration over editing: unlocking more effective image refinement!☆54May 26, 2026Updated 4 months ago
- CVPR 2021 Oral Paper PatchGenCN☆11Oct 28, 2021Updated 4 years ago
- Risky Object Localization (ROL) in a Driving Scene Dataset☆15Dec 24, 2023Updated 2 years ago
- Multi-Scale Attention for Audio Question Answering☆28Jul 19, 2023Updated 3 years ago