[ICIP 2022 oral] VLCap: Vision-Language with Contrastive Learning for Coherent Video Paragraph Captioning
☆28Jun 28, 2023Updated 3 years ago
Alternatives and similar repositories for VLCAP
Users that are interested in VLCAP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [Lab] lab website☆12Sep 3, 2026Updated 2 weeks ago
- [IJCV] AOE-Net: Entities Interactions Modeling with Adaptive Attention Mechanism for Temporal Action Proposals Generation☆20Jul 2, 2024Updated 2 years ago
- [WACV 2023] EmbryosFormer: Deformable Transformer and Collaborative Encoding-Decoding for Embryos Stage Development Classification☆26Apr 12, 2024Updated 2 years ago
- [ICRA 2024 Oral] Open-Fusion: Real-time Open-Vocabulary 3D Mapping and Queryable Scene Representation☆159Aug 19, 2024Updated 2 years ago
- [ISBI 2024] An implementation of TSRNet for ECG Anomaly Detection☆25Apr 11, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICPR 2022] 3DConvCaps: 3DUnet with Convolutional Capsule Encoder for Medical Image Segmentation☆48Jun 26, 2022Updated 4 years ago
- [IEEE BHI 2022] Multimodality Multi-Lead ECG Arrhythmia Classification using Self-Supervised Learning☆44Oct 4, 2024Updated last year
- [BMVC 2022] AISFormer: Amodal Instance Segmentation with Transformer☆47Nov 24, 2024Updated last year
- The repository of the ACCV 2024 paper "FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Ge…☆12Aug 15, 2026Updated last month
- [NeurIPS 2023] Official Implementation of "PaintSeg: Painting Pixels for Training-free Segmentation"☆14Dec 31, 2023Updated 2 years ago
- [Official Implementation] Acoustic Autoregressive Modeling 🔥☆74Aug 24, 2024Updated 2 years ago
- MELTR: Meta Loss Transformer for Learning to Fine-tune Video Foundation Models (CVPR 2023)☆35Apr 23, 2024Updated 2 years ago
- ☆13Jan 8, 2020Updated 6 years ago
- ☆10Sep 14, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆17Jun 4, 2026Updated 3 months ago
- ☆20Dec 4, 2025Updated 9 months ago
- Re-thinking Co-Salient Object Detection, TPAMI 2021☆24Jan 26, 2023Updated 3 years ago
- ☆10Oct 7, 2023Updated 2 years ago
- Supervision by Fusion: Towards Unsupervised Learning of Deep Salient Object Detector☆11Jun 24, 2023Updated 3 years ago
- Unofficial reimplementation of Dynamic Fusion with Intra- and Inter-modality Attention Flow for Visual Question Answering☆17Oct 30, 2019Updated 6 years ago
- Official PyTorch Implementation of paper EAN: Event Adaptive Network for Efficient Action Recognition https://arxiv.org/abs/2107.10771☆33Oct 24, 2023Updated 2 years ago
- Detectron2 is a platform for object detection, segmentation and other visual recognition tasks.☆19May 7, 2022Updated 4 years ago
- ☆14Jun 26, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for our IJCAI 2019 paper entitled "Conditional GAN with Discriminative Filter Generation for Text-to-Video Synthesis"☆14Mar 29, 2022Updated 4 years ago
- PyTorch implementation of soft-nms☆32May 23, 2026Updated 3 months ago
- Frozen Pretrained Transformers for Neural Sign Language Translation☆15Apr 23, 2022Updated 4 years ago
- Code and data for experiments on semantic fragments☆11Jun 23, 2022Updated 4 years ago
- EDUVSUM is a multimodal neural architecture that utilizes state-of-the-art audio, visual and textual features to identify important tempo…☆23Mar 8, 2024Updated 2 years ago
- A simple and effective feature extractor for untrimmed videos☆13Sep 1, 2022Updated 4 years ago
- [TPAMI'2023]Knowledge-enriched Attention Network with Group-wise Semantic for Visual Storytelling☆11Jan 3, 2023Updated 3 years ago
- ☆11Jun 27, 2023Updated 3 years ago
- ☆19May 15, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Experiments on GPT-3's ability to fit numerical models in-context.☆14Aug 11, 2022Updated 4 years ago
- ☆12May 26, 2023Updated 3 years ago
- [ACL 2020] PyTorch code for MART: Memory-Augmented Recurrent Transformer for Coherent Video Paragraph Captioning☆170Dec 4, 2020Updated 5 years ago
- [CVPR2022] Official code for Hierarchical Modular Network for Video Captioning. Our proposed HMN is implemented with PyTorch.☆50Sep 30, 2022Updated 3 years ago
- ☆11Sep 15, 2023Updated 3 years ago
- rendezvous-in-time☆15Sep 17, 2025Updated last year
- Official code for paper "Real-to-Sim for Highly Cluttered Environments via Physics-Consistent Inter-Object Reasoning"☆23May 16, 2026Updated 4 months ago