☆102Sep 19, 2024Updated last year
Alternatives and similar repositories for fineVideo
Users that are interested in fineVideo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A huge dataset for Document Visual Question Answering☆24Jul 29, 2024Updated 2 years ago
- Video-LlaVA fine-tune for CinePile evaluation☆51Aug 8, 2024Updated 2 years ago
- Competition of Mechanisms: Tracing How Language Models Handle Facts and Counterfactuals; ACL 2024☆13May 24, 2024Updated 2 years ago
- ANE accelerated embedding models!☆20Dec 11, 2024Updated last year
- 简单的手机推流服务器和RTSP服务器☆19Jun 4, 2018Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Unified Audio-Visual Perception for Multi-Task Video Localization☆33Apr 19, 2024Updated 2 years ago
- ☆23May 26, 2026Updated 2 months ago
- ☆17Oct 21, 2025Updated 9 months ago
- Unofficial Implementation of Selective Attention Transformer☆20Oct 31, 2024Updated last year
- YOLOv10: Real-Time End-to-End Object Detection☆12May 24, 2024Updated 2 years ago
- Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency☆62Jun 6, 2025Updated last year
- [ICLR 2025] AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark☆147Jun 4, 2025Updated last year
- My Blog.个人博客☆24May 15, 2018Updated 8 years ago
- ☆32Jul 29, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- This code is referring to a deep model that is used for micro-expression spotting with a CNN backbone and Transformer neck.☆23Aug 2, 2023Updated 3 years ago
- ☆22Jun 30, 2021Updated 5 years ago
- Official implement for LaserHuman.☆35Mar 29, 2025Updated last year
- 👾 E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding (NeurIPS 2024)☆74Jan 20, 2025Updated last year
- Hugging Face Jobs☆20Jul 11, 2025Updated last year
- Code used for the creation of OBELICS, an open, massive and curated collection of interleaved image-text web documents, containing 141M d…☆217Aug 28, 2024Updated last year
- [CVPR 2024] MovieChat: From Dense Token to Sparse Memory for Long Video Understanding☆706Jan 29, 2025Updated last year
- ☆25Dec 18, 2024Updated last year
- [ICML 2025] Official PyTorch implementation of LongVU☆432May 8, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆21Nov 18, 2024Updated last year
- [DEPRECIATED] [PyTorch 2.0] [638M] [85.33% acc] Full-attention multi-instrumental music transformer for supervised music generation, opti…☆33Nov 23, 2023Updated 2 years ago
- [PR 2024] A large Cross-Modal Video Retrieval Dataset with Reading Comprehension☆32Dec 28, 2023Updated 2 years ago
- [ECCV 2026] Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?☆97Jul 13, 2025Updated last year
- 合集☆35Apr 22, 2015Updated 11 years ago
- ☆94Apr 28, 2026Updated 3 months ago
- Official repository for "SSA: Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space"☆28May 7, 2026Updated 3 months ago
- Awesome papers & datasets specifically focused on long-term videos.☆382Oct 9, 2025Updated 10 months ago
- Video-R1: Reinforcing Video Reasoning in MLLMs [🔥the first paper to explore R1 for video]☆886Dec 14, 2025Updated 7 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- LinGPT, a GPT-4 webpage with just a single HTML file. 只有一个html文件的GPT4聊天网页,零门槛,10秒搞定。多Key轮询 Auto Key Rotation 支持代理平台/第三方Key Supports proxy…☆12Aug 28, 2023Updated 2 years ago
- ☆28Mar 3, 2025Updated last year
- VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning☆37May 9, 2026Updated 3 months ago
- Chineses-PPDB☆14Nov 23, 2020Updated 5 years ago
- Automatically derive Python dunder methods for your Rust code☆25May 26, 2026Updated 2 months ago
- ☆47Apr 9, 2025Updated last year
- 🔥🔥🔥 [IEEE TCSVT] Latest Papers, Codes and Datasets on Vid-LLMs.☆3,262Updated this week