Temporal Compact Bilinear Pooling (TCBP)
☆11May 27, 2020Updated 6 years ago
Alternatives and similar repositories for tcbp
Users that are interested in tcbp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the CVPR 2020 paper 'Action Modifiers: Learning from Adverbs in Instructional Videos'☆23May 17, 2021Updated 5 years ago
- ☆13May 10, 2025Updated last year
- A dataset for Audio-Visual Sound Event Detection in Movies☆26Jan 23, 2023Updated 3 years ago
- [arXiv 2020] Video Representation Learning with Visual Tempo Consistency☆24Jun 30, 2020Updated 6 years ago
- Just a data☆12Jun 25, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆11Dec 23, 2018Updated 7 years ago
- Official code of *Towards Event-oriented Long Video Understanding*☆12Jul 26, 2024Updated 2 years ago
- ☆12Jun 9, 2018Updated 8 years ago
- ☆32Jun 18, 2021Updated 5 years ago
- ComputeR vIsion for Sport Performance☆11May 14, 2024Updated 2 years ago
- Python program to generate, draw, and analyze spectral networks of class S theories☆12Mar 13, 2020Updated 6 years ago
- ☆10May 24, 2023Updated 3 years ago
- Code and datasets for "Text encoders are performance bottlenecks in contrastive vision-language models". Coming soon!☆11May 24, 2023Updated 3 years ago
- Code for the paper Joint Discovery of Object States and Manipulation Actions, ICCV 2017☆14Aug 7, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository contains the Adverbs in Recipes (AIR) dataset and the code published at the CVPR 23 paper: "Learning Action Changes by Me…☆13May 25, 2023Updated 3 years ago
- The code and data for "Summary-Oriented Vision Modeling for Multimodal Abstractive Summarization"☆11May 16, 2023Updated 3 years ago
- Evaluation script for VoxMovies dataset in PyTorch☆23Jan 12, 2024Updated 2 years ago
- ChangeIt dataset with more than 2600 hours of video with state-changing actions published at CVPR 2022☆11Mar 23, 2022Updated 4 years ago
- Implementation of "Slow-Fast Auditory Streams for Audio Recognition, ICASSP, 2021" in PyTorch☆73Sep 27, 2021Updated 4 years ago
- Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation[TNNLS2024]☆14May 6, 2025Updated last year
- ☆55Oct 16, 2023Updated 2 years ago
- ☆17Sep 25, 2023Updated 2 years ago
- Reading list for multimodal sequence learning☆14Sep 4, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆40May 7, 2022Updated 4 years ago
- SMILE: A Multimodal Dataset for Understanding Laughter☆13Jun 15, 2023Updated 3 years ago
- Official code of *Virgo: A Preliminary Exploration on Reproducing o1-like MLLM*☆20May 27, 2025Updated last year
- Code for "Distributed, Egocentric Representations of Graphs for Detecting Critical Structures" (ICML 2019)☆20Aug 24, 2021Updated 4 years ago
- Rotation equivariance meets local feature matching☆18Oct 20, 2022Updated 3 years ago
- ☆13Jul 20, 2024Updated 2 years ago
- A Jupyter Notebook and python scripts that allows users to easily train a siamese network on image similarity, export the model to a Save…☆21Oct 31, 2022Updated 3 years ago
- ☆19Jan 30, 2023Updated 3 years ago
- Latex template for Oxford integrated thesis☆20Apr 7, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- 🤖 基于A2A协议的智能合同风险审查系统 | AI-powered contract risk analysis system using A2A protocol☆15Jul 2, 2025Updated last year
- Former Research simulations and results☆10Jul 24, 2017Updated 9 years ago
- EACL 2023 paper "MLASK: Multimodal Summarization of Video-based News Articles"☆11Nov 7, 2023Updated 2 years ago
- tf&torch about nlp☆11Aug 12, 2022Updated 3 years ago
- All about FineGym (CVPR 2020 Oral): models, features, data, and more... keep starring and stay tuned!☆155Dec 26, 2024Updated last year
- Official code for the paper: "Metadata Archaeology"☆19May 10, 2023Updated 3 years ago
- YOLO-World-ONNX is a Python package for running inference on YOLO-WORLD Open-vocabulary-object detection model using ONNX models. It prov…☆17Feb 6, 2026Updated 5 months ago