Code for "Visual Spatial Description: Controlled Spatial-Oriented Image-to-Text Generation"
☆25Mar 9, 2024Updated 2 years ago
Alternatives and similar repositories for VSD
Users that are interested in VSD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Apr 3, 2026Updated 3 months ago
- [ECCV2022] A PyTorch implementation of the paper "Spatial and Visual Perspective-Taking via View Rotation and Relation Reasoning for Embo…☆13Mar 20, 2023Updated 3 years ago
- [ECCV2024] Nonverbal Interaction Detection☆31Oct 30, 2024Updated last year
- Collection of evaluation code for natural language generation.☆12Jan 6, 2021Updated 5 years ago
- Data preprocessing for IUPUI-CSRC Pedestrian Situated Intent (PSI) benchmark dataset.☆11Oct 5, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Lyra: A Benchmark for Turducken-Style Code Generation☆15Apr 22, 2022Updated 4 years ago
- ☆28Mar 20, 2023Updated 3 years ago
- This repositary hosts my experiments for the project, I did with OffNote Labs.☆10Apr 12, 2021Updated 5 years ago
- DOZE: A Dataset for Open-Vocabulary Zero-Shot Object Navigation in Dynamic Environments☆25Apr 8, 2025Updated last year
- Official implementation of the WACV 2023 paper "Benchmarking Visual Localization for Autonomous Navigation".☆24Sep 25, 2023Updated 2 years ago
- Monitor Google Scholar author citation counts and track changes automatically without opening tabs.☆73Jul 22, 2026Updated last week
- The implementation of Text Classification with Negative Supervision (ACL, 2020)☆10Oct 8, 2020Updated 5 years ago
- 利用BERT预训练模型进行文本生成,可用于对话、摘要、问题生成等任务。 目前支持策略,词表的插入和删除、自定义Character Embedding、随机词替换等☆10Jun 1, 2022Updated 4 years ago
- Code of the ICCV 2023 paper "March in Chat: Interactive Prompting for Remote Embodied Referring Expression"☆26May 22, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆13Jul 20, 2022Updated 4 years ago
- Code for the paper : "Weakly-supervised learning of visual relations", ICCV17☆40Oct 20, 2017Updated 8 years ago
- Prompt-Guided Retrieval For Non-Knowledge-Intensive Tasks☆12Sep 1, 2023Updated 2 years ago
- ☆16Mar 17, 2025Updated last year
- IsoBN: Fine-Tuning BERT with Isotropic Batch Normalization☆12Nov 23, 2021Updated 4 years ago
- ☆11Jan 3, 2023Updated 3 years ago
- ☆14Jan 10, 2024Updated 2 years ago
- The code and data for "Summary-Oriented Vision Modeling for Multimodal Abstractive Summarization"☆11May 16, 2023Updated 3 years ago
- GRAIN: Gradient-based Intra-attention Pruning on Pre-trained Language Models☆19Jul 12, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Basic exercises of chinese information processing☆36Sep 1, 2021Updated 4 years ago
- ☆18Sep 19, 2025Updated 10 months ago
- MXNet复现SSD目标检测网络☆12Apr 2, 2019Updated 7 years ago
- Visual Question Generation☆11Aug 20, 2024Updated last year
- Here is the repo for public scripts.☆12Jul 16, 2022Updated 4 years ago
- [CVPR 2025] LoRA Recycle: Unlocking Tuning-Free Few-Shot Adaptability in Visual Foundation Models by Recycling Pre-Tuned LoRAs☆14Jun 20, 2025Updated last year
- [AAAI 2023] Official implementation of FiTs: Fine-grained Two-stage Training for Knowledge Base Question Answering☆11Mar 10, 2023Updated 3 years ago
- Cross-Perspective Topic Modeling☆11Oct 27, 2017Updated 8 years ago
- Code for "Inducer-tuning: Connecting Prefix-tuning and Adapter-tuning" (EMNLP 2022) and "Empowering Parameter-Efficient Transfer Learning…☆11Feb 6, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- FIGR-8, but images in .SVG vector graphics format☆15Feb 16, 2019Updated 7 years ago
- Official code for our AAAI25 oral👑 paper Harmonious Group Choreography with Trajectory-Controllable Diffusion — hope you enjoy exploring…☆20Oct 3, 2025Updated 9 months ago
- ☆10Jul 24, 2023Updated 3 years ago
- Boundaries and Region Representation Fusion☆12Mar 24, 2023Updated 3 years ago
- [ICCV 2023] With a Little Help from your own Past: Prototypical Memory Networks for Image Captioning.☆19Jun 7, 2024Updated 2 years ago
- 自己阅读的多模态对话系统论文(及部分笔记)汇总☆22Jan 5, 2023Updated 3 years ago
- [KDD'23] This is the code repo for our KDD'23 paper "DyGen: Learning from Noisy Labels via Dynamics-Enhanced Generative Modeling".☆11Jun 14, 2023Updated 3 years ago