Offical Code for GPT4Video: A Unified Multimodal Large Language Model for lnstruction-Followed Understanding and Safety-Aware Generation
☆144Oct 30, 2024Updated last year
Alternatives and similar repositories for GPT4Video
Users that are interested in GPT4Video are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Cross Sentence Neural Machine Translation☆10Mar 26, 2018Updated 8 years ago
- ☆20Jun 17, 2024Updated 2 years ago
- TVsub: DCU-Tencent Chinese-English Dialogue Corpus☆47Feb 14, 2018Updated 8 years ago
- ☆18Dec 18, 2023Updated 2 years ago
- GPT 4 Vision + TTS 多模态能力 Demo☆17Nov 15, 2023Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Official code for Goldfish model for long video understanding and MiniGPT4-video for short video understanding☆638Dec 10, 2024Updated last year
- ACL'24 (Oral) Tuning Large Multimodal Models for Videos using Reinforcement Learning from AI Feedback☆76Sep 12, 2024Updated last year
- ☆21Oct 10, 2023Updated 2 years ago
- ☆32Jan 25, 2024Updated 2 years ago
- ☆25Dec 21, 2023Updated 2 years ago
- Feature Decay Algorithms☆11Mar 5, 2014Updated 12 years ago
- The benchmark and datasets of the ICML 2024 paper "VisionGraph: Leveraging Large Multimodal Models for Graph Theory Problems in Visual C…☆17May 27, 2024Updated 2 years ago
- ☆16Sep 6, 2024Updated last year
- ☆18Dec 29, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [NeurIPS 2023 Datasets and Benchmarks] "FETV: A Benchmark for Fine-Grained Evaluation of Open-Domain Text-to-Video Generation", Yuanxin L…☆56Mar 4, 2024Updated 2 years ago
- ☆19Jan 15, 2024Updated 2 years ago
- ☆17Jan 10, 2024Updated 2 years ago
- ☆81Nov 24, 2024Updated last year
- ☆13Dec 18, 2023Updated 2 years ago
- ☆159Oct 31, 2024Updated last year
- ☆70Mar 3, 2024Updated 2 years ago
- Multilingual Corpus of Web Fiction☆205Jun 28, 2024Updated 2 years ago
- ☆14Dec 26, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- LLaMA-VID: An Image is Worth 2 Tokens in Large Language Models (ECCV 2024)☆862Jul 29, 2024Updated 2 years ago
- ☆30Dec 19, 2023Updated 2 years ago
- Official code for the paper "Does CLIP's Generalization Performance Mainly Stem from High Train-Test Similarity?" (ICLR 2024)☆11Aug 26, 2024Updated 2 years ago
- [CVPR 2024] Official PyTorch implementation of the paper "One For All: Video Conversation is Feasible Without Video Instruction Tuning"☆35Feb 2, 2024Updated 2 years ago
- ☆134Feb 13, 2024Updated 2 years ago
- Project for SNARE benchmark☆11Jun 5, 2024Updated 2 years ago
- ☆34Jan 16, 2024Updated 2 years ago
- [EMNLP 2025 Findings] MEXA: Towards General Multimodal Reasoning with Dynamic Multi-Expert Aggregation☆15Aug 22, 2025Updated last year
- Official pytorch implementation for SingleInsert☆28Apr 19, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆16Nov 28, 2023Updated 2 years ago
- ☆11Aug 28, 2023Updated 3 years ago
- OpenVLThinker [NeurIPS 2025] & OpenVLThinkerV2 [COLM 2026]☆156May 25, 2026Updated 3 months ago
- Repository for evaluating Pegasus-1 and video-language foundation models☆14Nov 12, 2024Updated last year
- ☆13Oct 12, 2023Updated 2 years ago
- ☆13Feb 28, 2024Updated 2 years ago
- 【EMNLP 2024🔥】Video-LLaVA: Learning United Visual Representation by Alignment Before Projection☆3,500Dec 3, 2024Updated last year