Pytorch implementation of Twelve Labs' Video Foundation Model evaluation framework & open embeddings
☆36Aug 23, 2024Updated 2 years ago
Alternatives and similar repositories for video-embeddings-evaluation-framework
Users that are interested in video-embeddings-evaluation-framework are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Study Friendly Implementation of Pix2Pix in Tensorflow☆13Sep 8, 2018Updated 8 years ago
- [WACV 2026] MomentMix Augmentation with Length-Aware DETR for Temporally Robust Moment Retrieval☆15Sep 18, 2025Updated last year
- [NeurIPS 2024] Activating Self-Attention for Multi-Scene Absolute Pose Regression☆15Feb 24, 2025Updated last year
- Official PyTorch implementation of "Threshold Matters in WSSS: Manipulating the Activation for the Robust and Accurate Segmentation Model…☆40Jul 10, 2022Updated 4 years ago
- Expanded Adaptive Scaling Normalization for End to End Image Compression☆11Sep 4, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official Implementation of "PsyNet: Self-supervised Approach to Object Localization Using Point Symmetric Transformation"☆25Dec 8, 2022Updated 3 years ago
- Commonality in Natural Images Rescues GANs: Pretraining GANs with Generic and Privacy-free Synthetic Data - Official PyTorch Implementati…☆34Nov 14, 2022Updated 3 years ago
- A lightweight plugin that integrates TwelveLabs video understanding capabilities directly into Claude Code - enabling semantic video sear…☆23May 14, 2026Updated 4 months ago
- MAtch, eXpand and Improve: Unsupervised Finetuning for Zero-Shot Action Recognition with Language Knowledge (ICCV 2023)☆31Sep 5, 2023Updated 3 years ago
- Accelerating the development of large multimodal models (LMMs) with lmms-eval☆14Oct 14, 2024Updated last year
- Pytorch implementation of HyperLLaVA: Dynamic Visual and Language Expert Tuning for Multimodal Large Language Models☆28Mar 22, 2024Updated 2 years ago
- Downscaling Intelligence: Exploring Perception and Reasoning Bottlenecks in Small Multimodal Models☆26Mar 21, 2026Updated 6 months ago
- various tools to download, convert and process the full text of scientific articles☆10Apr 2, 2024Updated 2 years ago
- ☆34May 14, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Official PyTorch implementation of “MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation”☆18Dec 5, 2024Updated last year
- Butteraugli estimating filter for the psychovisual similarity of two images.☆13Jul 23, 2023Updated 3 years ago
- A Vision-Language Benchmark for Microscopy Understanding☆31Mar 13, 2025Updated last year
- ☆81Nov 24, 2024Updated last year
- ☆11Dec 2, 2024Updated last year
- Code for EMNLP25 paper "Video-RTS: Rethinking Reinforcement Learning and Test-Time Scaling for Efficient and Enhanced Video Reasoning"☆24Feb 18, 2026Updated 7 months ago
- Concise Reasoning via Reinforcement Learning☆13Apr 16, 2025Updated last year
- Create embeddings for LLM using the Nomic API☆23Nov 21, 2024Updated last year
- Addon to automate blender baking process☆13Feb 13, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code and datasets for "Text encoders are performance bottlenecks in contrastive vision-language models". Coming soon!☆11May 24, 2023Updated 3 years ago
- [CVPR2023] Masked Video Distillation: Rethinking Masked Feature Modeling for Self-supervised Video Representation Learning (https://arxiv…☆137May 21, 2023Updated 3 years ago
- ☆18Feb 13, 2025Updated last year
- This repository contains the Adverbs in Recipes (AIR) dataset and the code published at the CVPR 23 paper: "Learning Action Changes by Me…☆13May 25, 2023Updated 3 years ago
- [ECCV 2024] Elysium: Exploring Object-level Perception in Videos via MLLM☆89Oct 25, 2024Updated last year
- [CVPR23 Highlight] CREPE: Can Vision-Language Foundation Models Reason Compositionally?☆36Apr 27, 2023Updated 3 years ago
- ☆13Feb 21, 2024Updated 2 years ago
- Code used to run experiments for the ICLR 2023 paper "Computational Language Acquisition with Theory of Mind".☆15Apr 27, 2023Updated 3 years ago
- [CVPR 2023] Exploring Discontinuity for Video Frame Interpolation Official, Highlight☆37May 23, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2023 Spotlight] Code for "Contrastive Lift: 3D Object Instance Segmentation by Slow-Fast Contrastive Fusion"☆73Nov 3, 2023Updated 2 years ago
- Official implementation of "HowToCaption: Prompting LLMs to Transform Video Annotations at Scale." ECCV 2024☆60Aug 19, 2025Updated last year
- Pragmatic models for generating and following instructions☆13Dec 22, 2019Updated 6 years ago
- Web Interface for gaze recording: CVPR 2018☆10Jul 10, 2018Updated 8 years ago
- Official implementation of the paper "STARS: Self-supervised 3D Action Recognition with Contrastive Tuning".☆18Jan 6, 2025Updated last year
- 4D Strategic Memory Engine for Autonomous AI Agents — store wisdom, not just facts☆16Apr 3, 2026Updated 5 months ago
- SMILE: A Multimodal Dataset for Understanding Laughter☆13Jun 15, 2023Updated 3 years ago