The official code of "CSTA: CNN-based Spatiotemporal Attention for Video Summarization"
☆71Jul 27, 2025Updated last year
Alternatives and similar repositories for CSTA
Users that are interested in CSTA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pytorch implementation for "Progressive Video Summarization via Multimodal Self-supervised Learning"☆35Aug 26, 2025Updated last year
- The official implementation of 'Align and Attend: Multimodal Summarization with Dual Contrastive Losses' (CVPR 2023)☆86Apr 24, 2023Updated 3 years ago
- Deep learning model for supervised video summarization called Multi Source Visual Attention (MSVA)☆47Mar 21, 2024Updated 2 years ago
- Source code for the paper "Unsupervised Video Summarization via Multi-source Features" published at ICMR 2021☆21Apr 5, 2022Updated 4 years ago
- A PyTorch Implementation of PGL-SUM from "Combining Global and Local Attention with Positional Encoding for Video Summarization" (IEEE IS…☆93Jan 30, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Pytorch code for paper Contrastive Losses Are Natural Criteria for Unsupervised Video Summarization☆22Jan 7, 2023Updated 3 years ago
- [TMM 2023] VideoXum: Cross-modal Visual and Textural Summarization of Videos☆53Apr 9, 2024Updated 2 years ago
- DSNet: A Flexible Detect-to-Summarize Network for Video Summarization☆223Sep 16, 2021Updated 4 years ago
- A PyTorch implementation of the software used in: "A study on the use of attention for explaining video summarization" (NarSUM Workshop a…☆11Oct 20, 2023Updated 2 years ago
- This is the implementation of the paper Video Summarization by Learning from Unpaired Data(CVPR2019)☆37Sep 5, 2019Updated 7 years ago
- Simple video summarisation Python package.☆25Jan 29, 2024Updated 2 years ago
- [ECCV'24 Workshops Oral] DALDA: Data Augmentation Leveraging Diffusion Model and LLM with Adaptive Guidance Scaling☆33Feb 6, 2026Updated 7 months ago
- Official Code for paper "Towards Efficient and Effective Unlearning of Large Language Models for Recommendation" (Frontiers of Computer S…☆37Jul 19, 2024Updated 2 years ago
- Implementation of LTC-SUM: Lightweight Client-driven Personalized Video Summarization Framework Using 2D CNN☆22Jul 11, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Evaluation and dataset construction code for the CVPR 2025 paper "Vision-Language Models Do Not Understand Negation"☆49Feb 26, 2026Updated 6 months ago
- Official code and dataset link for ''VMSMO: Learning to Generate Multimodal Summary for Video-based News Articles''☆36Jul 30, 2021Updated 5 years ago
- This is an official PyTorch Implementation of Neighbor Relations Matter in Video Scene Detection.☆30Mar 19, 2025Updated last year
- [CVPR'25] Official implementation of the paper "Not Only Text: Exploring Compositionality of Visual Representations in Vision-Language Mo…☆18Nov 21, 2025Updated 9 months ago
- Official pytorch repository for CG-DETR "Correlation-guided Query-Dependency Calibration in Video Representation Learning for Temporal Gr…☆155Aug 21, 2024Updated 2 years ago
- A Video Summarization framework for implementation and benchmark of Deep Learning models☆33Sep 9, 2024Updated last year
- [ACMMM 2024] An Inverse Partial Optimal Transport Framework for Music-guided Movie Trailer Generation☆17Mar 15, 2025Updated last year
- MICCAI 2021 Code for the paper: Ultrasound Video Transformers for Cardiac Ejection Fraction Estimation☆40Nov 19, 2021Updated 4 years ago
- Code for paper, "TL;DW? Summarizing Instructional Videos with Task Relevance & Cross-Modal Saliency" ECCV 2022☆39Feb 17, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆57Nov 1, 2024Updated last year
- source code of our MGPN in SIGIR 2022☆18Jun 8, 2022Updated 4 years ago
- Visual Dependency Transformers: Dependency Tree Emerges from Reversed Attention (CVPR 2023)☆32Mar 28, 2023Updated 3 years ago
- ☆27Jan 4, 2025Updated last year
- This code is a python implementation of the paper, "Illumination Estimation for Nature Preserving Low Light Image Enhancement",in 2020.☆12Jan 12, 2021Updated 5 years ago
- [WACV 2025] Official Pytorch code for "Background-aware Moment Detection for Video Moment Retrieval"☆16Feb 24, 2025Updated last year
- EACL 2023 paper "MLASK: Multimodal Summarization of Video-based News Articles"☆11Nov 7, 2023Updated 2 years ago
- A Tensorflow implementation of Speech Emotion Recognition using Audio signals and Text Data☆12May 16, 2022Updated 4 years ago
- State-Relabeling Adversarial Active Learning☆14Aug 17, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Video Feature Extraction Code for EMNLP 2020 paper "HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training"☆118Jun 9, 2021Updated 5 years ago
- [CVPR 2024] MMSum: A Dataset for Multimodal Summarization and Thumbnail Generation of Videos☆38Jan 29, 2025Updated last year
- ☆11Jul 11, 2025Updated last year
- Code for the paper "Unbiased Supervised Contrastive Learning" | ICLR 2023 https://openreview.net/forum?id=Ph5cJSfD2XN☆12Sep 22, 2023Updated 2 years ago
- ☆14Feb 26, 2024Updated 2 years ago
- [IEEE OJSP'26, IEEE SLT'24] "Speaker-Disentangled Chunk-Wise Regression for Syllabic Tokenization"☆46Updated this week
- Learning to cut end-to-end pretrained modules☆38Apr 17, 2025Updated last year