[AAAI 2025] 🎬RCDMs🎬: Boosting Consistency in Story Visualization with Rich-Contextual Conditional Diffusion Models. RCDMs improve story generation with strong semantic and temporal consistency, integrating rich contextual conditions and enabling one-pass inference for enhanced coherence.
☆64Sep 30, 2025Updated 11 months ago
Alternatives and similar repositories for RCDMs
Users that are interested in RCDMs are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI-2025] Official Codes for “Motif Guided Graph Transformer with Combinatorial Skeleton Prototype Learning for Skeleton-Based Person R…☆21Feb 26, 2025Updated last year
- This is the official code implement for AAAI 2025 paper ``Defeasible Visual Entailment: Benchmark, Evaluator, and Reward-Driven Optimizat…☆22Mar 21, 2025Updated last year
- [AAAI 2025] Detecting and Mitigating Hallucination in Large Vision Language Models via Fine-Grained AI Feedback☆36Dec 16, 2025Updated 9 months ago
- AAAI '25. Retrieval-Augmented Multimodal Social Media Popularity Prediction☆26Sep 15, 2026Updated last week
- [AAAI 2025] CoPEFT: Fast Adaptation Framework for Multi-Agent Collaborative Perception with Parameter-Efficient Fine-Tuning☆28Apr 14, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [AAAI 2025] Official implementation of the paper "Exploring Semantic Consistency and Style Diversity for Domain Generalized Semantic Segm…☆43Dec 17, 2024Updated last year
- [AAAI 2025] RhythmMamba☆114Jul 29, 2025Updated last year
- [AAAI2025] FedCFA: Alleviating Simpson’s Paradox in Model Aggregation with Counterfactual Federated Learning☆24Jan 23, 2025Updated last year
- [AAAI 2025] Depth-Centric Dehazing and Depth-Estimation from Real-World Hazy Driving Video☆119Jul 25, 2026Updated last month
- ☆71Dec 18, 2024Updated last year
- 【AAAI2025】MambaPro: Multi-Modal Object Re-Identification with Mamba Aggregation and Synergistic Prompt☆93May 13, 2025Updated last year
- AAAI 2025: Hierarchical Consensus Network for Multiview Feature Learning☆18Feb 5, 2025Updated last year
- 【AAAI2025】DeMo: Decoupled Feature-Based Mixture of Experts for Multi-Modal Object Re-Identification☆77Mar 9, 2025Updated last year
- ☆59Jun 14, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official implementation of the paper "Attentive Eraser: Unleashing Diffusion Model’s Object Removal Potential via Self-Attention Redirect…☆221Jun 16, 2026Updated 3 months ago
- AAAI 2025: Autonomous LLM-enhanced adversarial attack for text-to-motion☆19Sep 15, 2025Updated last year
- 【CVPR2025】IDEA: Inverted Text with Cooperative Deformable Aggregation for Multi-modal Object Re-Identification☆54Apr 8, 2025Updated last year
- [NeurIPS 2024] 🕺IMAGPose🕺: A Unified Conditional Framework for Pose-Guided Person Generation. IMAGPose enables versatile pose-guided im…☆169Sep 30, 2025Updated 11 months ago
- [2025] Language-driven Motion Prior Knowledge Learning for Moving Infrared Small Target Detection☆50Jun 17, 2026Updated 3 months ago
- Official code for CustAny: Customizing Anything from A Single Example. Accepted by CVPR2025 (Oral)☆48Apr 10, 2025Updated last year
- This repo is the official implementation of "Retrieval-Augmented Dynamic Prompt Tuning for Incomplete Multimodal Learning" accepted by AA…☆67May 26, 2026Updated 3 months ago
- TheaterGen: Character Management with LLM for Consistent Multi-turn Image Generation☆70Sep 26, 2024Updated last year
- ☆27Mar 16, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ECCV2024] StoryImager: A Unified and Efficient Framework for Coherent Story Visualization and Completion☆40Jul 5, 2024Updated 2 years ago
- [AAAI 2025] PAT: Pruning-Aware Tuning for Large Language Models☆38Feb 1, 2025Updated last year
- Implementation code:Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models☆192Sep 30, 2025Updated 11 months ago
- AugTarget data augmentation for infrared small target detection.☆20May 19, 2023Updated 3 years ago
- [ECCV24] Attention Regulation on T2I Diffusion Models☆19Jul 8, 2024Updated 2 years ago
- [AAAI 2025] Official implementation of the paper "EOV-Seg: Efficient Open-Vocabulary Panoptic Segmentation"☆40Dec 17, 2024Updated last year
- [TGRS 2024] SSTNet: Sliced spatio-temporal network with cross-slice ConvLSTM for moving infrared dim-small target detection☆68Jan 21, 2025Updated last year
- [AAAI2025] Official implementation of the paper "RAP-SR: RestorAtion Prior Enhancement in Diffusion Models for Realistic Image Super-Reso…☆17Mar 22, 2025Updated last year
- ☆17Jul 23, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- LiDAR-NeRF: Novel LiDAR View Synthesis via Neural Radiance Fields☆163Jul 17, 2023Updated 3 years ago
- [WACV 2025 oral] All-in-One Image Compression and Restoration.☆35Apr 5, 2026Updated 5 months ago
- [ICLR 2023] Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models☆50Jan 17, 2024Updated 2 years ago
- Prototypical Contrast and Reverse Prediction: Unsupervised Skeleton based Action Recognition☆11Aug 30, 2021Updated 5 years ago
- VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis☆125Mar 25, 2026Updated 5 months ago
- ☆22Apr 17, 2024Updated 2 years ago
- CutDiffusion: A Simple, Fast, Cheap, and Strong Diffusion Extrapolation Method☆27Oct 9, 2025Updated 11 months ago