FunQA benchmarks funny, creative, and magic videos for challenging tasks including timestamp localization, video description, reasoning, and beyond.
☆104Dec 25, 2025Updated 7 months ago
Alternatives and similar repositories for FunQA
Users that are interested in FunQA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Syphus: Automatic Instruction-Response Generation Pipeline☆14Dec 14, 2023Updated 2 years ago
- Benchmarking and Analyzing Generative Data for Visual Recognition☆26Jul 25, 2023Updated 3 years ago
- Relate Anything Model is capable of taking an image as input and utilizing SAM to identify the corresponding mask within the image.☆473Jul 4, 2023Updated 3 years ago
- A local AI assistant running on your device. It turns your files into actionable memory.☆58Mar 24, 2026Updated 4 months ago
- On-Device Domain Generalization☆47Nov 9, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code and datasets for "Text encoders are performance bottlenecks in contrastive vision-language models". Coming soon!☆11May 24, 2023Updated 3 years ago
- [ECCV 2022] StyleLight: HDR Panorama Generation for Lighting Estimation and Editing☆149Oct 9, 2023Updated 2 years ago
- [IJCV 2026] Official Code for "PointHPS: Cascaded 3D Human Pose and Shape Estimation from Point Clouds"☆71Feb 12, 2026Updated 6 months ago
- [CVPR 2022 Oral] Versatile Multi-Modal Pre-Training for Human-Centric Perception☆125Jun 23, 2022Updated 4 years ago
- General video interaction platform based on LLMs, including Video ChatGPT☆257Jul 26, 2023Updated 3 years ago
- 🦦 Otter, a multi-modal model based on OpenFlamingo (open-sourced version of DeepMind's Flamingo), trained on MIMIC-IT and showcasing imp…☆3,432Mar 5, 2024Updated 2 years ago
- Toolbox for HuMMan Dataset☆128Dec 7, 2024Updated last year
- [NeurIPS2023] Official implementation of the paper "Large Language Models are Visual Reasoning Coordinators"☆106Nov 9, 2023Updated 2 years ago
- Official Code for "Digital Life Project: Autonomous 3D Characters with Social Intelligence"☆42Sep 9, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [NeurIPS 2021] Garment4D: Garment Reconstruction from Point Cloud Sequences☆138Dec 11, 2021Updated 4 years ago
- A framework that allows you to apply Sparse AutoEncoder on any models☆53Jul 11, 2025Updated last year
- [TPAMI] Searching prompt modules for parameter-efficient transfer learning.☆241Dec 8, 2023Updated 2 years ago
- ☆79May 4, 2025Updated last year
- Code for our IJCV paper "HumanLiff: Layer-wise 3D Human Generation with Diffusion Model"☆55Apr 11, 2026Updated 4 months ago
- [NeurIPS 2023] Official Code for "Towards Robust and Expressive Whole-body Human Pose and Shape Estimation"☆50Feb 13, 2026Updated 5 months ago
- [ECCV2022] New benchmark for evaluating pre-trained model; New supervised contrastive learning framework.☆110Dec 8, 2023Updated 2 years ago
- Privacy-first AI memory layer - Signal for AI Memory. E2EE, local-first, works with Claude, Cursor, and any MCP-compatible AI.☆23Jun 12, 2026Updated 2 months ago
- ConsistentNeRF Enhances Neural Radiance Fields with 3D Consistency for Sparse View Synthesis☆75Oct 12, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [IJCV 2025] Code for DeepFake-Adapter: Dual-Level Adapter for DeepFake Detection☆63Dec 24, 2024Updated last year
- Long Context Transfer from Language to Vision☆408Mar 18, 2025Updated last year
- [NeurIPS 2021] ORL: Unsupervised Object-Level Representation Learning from Scene Images☆58Dec 6, 2021Updated 4 years ago
- [ECCV 2022 & IJCV 2025] PyTorch code for SeqDeepFake: Detecting and Recovering Sequential DeepFake Manipulation☆151Dec 3, 2024Updated last year
- [ICCV 2025] Auto Interpretation Pipeline and many other functionalities for Multimodal SAE Analysis.☆199Sep 26, 2025Updated 10 months ago
- The official repository of "Video assistant towards large language model makes everything easy"☆232Dec 24, 2024Updated last year
- [IEEE TPAMI-2024] Pair then Relation: Pair-Net for Panoptic Scene Graph Generation☆101Nov 20, 2024Updated last year
- [ICML 2025] Streamline Without Sacrifice - Squeeze out Computation Redundancy in LMM☆20May 22, 2025Updated last year
- Official code release for DeformToon3D: Deformable 3D Toonification from Neural Radiance Fields (ICCV 2023)☆56May 24, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- An official codebase for "NormLens: Reading Books is Great, But Not if You Are Driving! Visually Grounded Reasoning about Defeasible Comm…☆10May 9, 2024Updated 2 years ago
- [CVPR2023] All in One: Exploring Unified Video-Language Pre-training☆281Mar 25, 2023Updated 3 years ago
- ☆27Oct 5, 2023Updated 2 years ago
- ☆78Apr 9, 2026Updated 4 months ago
- NExT-QA: Next Phase of Question-Answering to Explaining Temporal Actions (CVPR'21)☆189Aug 2, 2025Updated last year
- ☆159Oct 31, 2024Updated last year
- Official Implementation of ICCV 2023 paper "StyleInV: A Temporal Style Modulated Inversion Network for Unconditional Video Generation"☆23May 10, 2024Updated 2 years ago