☆34Jun 2, 2023Updated 3 years ago
Alternatives and similar repositories for TranS4mer
Users that are interested in TranS4mer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆143Jan 3, 2024Updated 2 years ago
- ☆58Dec 2, 2025Updated 9 months ago
- Learning Interactions and Relationships between Movie Characters (CVPR'20)☆22Apr 12, 2023Updated 3 years ago
- Code for CVPR 2022 paper "Scene Consistency Representation Learning for Video Scene Segmentation"☆113Feb 14, 2023Updated 3 years ago
- [ACL 2023] VSTAR is a multimodal dialogue dataset with scene and topic transition information☆16Oct 27, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆21Mar 22, 2023Updated 3 years ago
- Map (deep learning) model weights between different model implementations.☆19Mar 9, 2026Updated 6 months ago
- ☆11May 12, 2024Updated 2 years ago
- ☆24Sep 24, 2023Updated 3 years ago
- Multi-modal transformer approach for natural language query based joint video summarization and highlight detection☆17May 23, 2024Updated 2 years ago
- Official PyTorch implementation of "DiGA: Distil to Generalize and then Adapt for Domain Adaptive Semantic Segmentation" (CVPR 2023)☆29Apr 1, 2024Updated 2 years ago
- Code for CVPR2023 paper "Collaborative Noisy Label Cleaner: Learning Scene-aware Trailers for Multi-modal Highlight Detection in Movies"☆18Mar 21, 2023Updated 3 years ago
- A new multi-shot video understanding benchmark Shot2Story with comprehensive video summaries and detailed shot-level captions.☆181Jan 30, 2025Updated last year
- Character-aware audio-only subtitling☆31Jun 15, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [Interspeech 2026] Revisiting Active Speaker Detection: An In-the-Wild Benchmark for Generalization and Robustness☆24Jun 25, 2026Updated 3 months ago
- ☆19Aug 13, 2024Updated 2 years ago
- Repo from the "Learning with limited labeled data" seminar @ Uni of Tuebingen. A collection of notes, notebooks and slideshows to underst…☆17Apr 13, 2023Updated 3 years ago
- Code for the paper: Graph Jigsaw Learning for Cartoon Face Recognition☆10Jul 1, 2022Updated 4 years ago
- Provably (and non-vacuously) bounding test error of deep neural networks under distribution shift with unlabeled test data.☆10Feb 27, 2024Updated 2 years ago
- Compare NVIDIA Video Codec SDK's, PyAV's, and OpenCV's performance on video decoding.☆13Dec 18, 2022Updated 3 years ago
- A custom Open WebUI Pipe for LangGraph with real-time human-in-the-loop control.☆16Nov 4, 2025Updated 10 months ago
- This repository contains the codebase for MovieCLIP: Visual Scene Recognition in Movies☆43Oct 1, 2023Updated 2 years ago
- JoinAI是一个开源仓库,专注于算法工程能力的培养,包括工程和数学原理的整理☆11Apr 20, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 为视障人群生成电影,输入是电影剧本和mkv格式电影,输出为带有解说的电影☆12Jul 28, 2019Updated 7 years ago
- [ICML 2024] Official code for Uncertainty Estimation by Density Aware Evidential Deep Learning☆17Mar 20, 2026Updated 6 months ago
- The implementation code of AAAI 2020 paper "Pixel-aware Deep Function-mixture Network for Spectral Super-Resolution".☆12Dec 31, 2019Updated 6 years ago
- FBNet code for FGOC in aerial images☆15Jun 9, 2022Updated 4 years ago
- Demo code for 'Unsupervised and Unregistered Hyperspectral Image Super-Resolution with Mutual Dirichlet-Net'.☆13Mar 14, 2023Updated 3 years ago
- ☆16Aug 13, 2024Updated 2 years ago
- ☆85Mar 10, 2025Updated last year
- A RESTful API for Microsoft's TRELLIS, enabling easy evaluation of machine learning model robustness. Features include model upload, robu…☆17Dec 17, 2024Updated last year
- Quick Long Video Understanding [TMLR2025]☆78Oct 27, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Feb 18, 2024Updated 2 years ago
- [NeurIPS 2023 D&B] VidChapters-7M: Video Chapters at Scale☆214Nov 13, 2023Updated 2 years ago
- ☆10Jun 5, 2023Updated 3 years ago
- Github repo for referring atomic video action recognition☆20Oct 2, 2024Updated last year
- A series of Jupyter notebooks that walk you through the fundamentals of Machine Learning and Deep Learning in python using Scikit-Learn a…☆11Jun 8, 2019Updated 7 years ago
- 【CVPR'24】OST: Refining Text Knowledge with Optimal Spatio-Temporal Descriptor for General Video Recognition☆39Apr 27, 2024Updated 2 years ago
- Generate interleaved text and image content in a structured format you can directly pass to downstream APIs.☆29Oct 18, 2024Updated last year