[AAAI 2025] 🎬RCDMs🎬: Boosting Consistency in Story Visualization with Rich-Contextual Conditional Diffusion Models. RCDMs improve story generation with strong semantic and temporal consistency, integrating rich contextual conditions and enabling one-pass inference for enhanced coherence.
☆64Sep 30, 2025Updated 11 months ago
Alternatives and similar repositories for RCDMs
Users that are interested in RCDMs are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI-2025] Official Codes for “Motif Guided Graph Transformer with Combinatorial Skeleton Prototype Learning for Skeleton-Based Person R…☆21Feb 26, 2025Updated last year
- Source code for AAAI'25 paper "Component-Level Segmentation for Oracle Bone Inscription Decipherment"☆20Oct 13, 2025Updated 10 months ago
- This is the official code implement for AAAI 2025 paper ``Defeasible Visual Entailment: Benchmark, Evaluator, and Reward-Driven Optimizat…☆22Mar 21, 2025Updated last year
- [AAAI 2025] Detecting and Mitigating Hallucination in Large Vision Language Models via Fine-Grained AI Feedback☆36Dec 16, 2025Updated 8 months ago
- AAAI '25. Retrieval-Augmented Multimodal Social Media Popularity Prediction☆24Jul 8, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [AAAI 2025] CoPEFT: Fast Adaptation Framework for Multi-Agent Collaborative Perception with Parameter-Efficient Fine-Tuning☆28Apr 14, 2025Updated last year
- ✨ [AAAI 2025] Queryable Prototype Multiple Instance Learning with Vision-Language Models for Incremental Whole Slide Image Classification☆54Apr 16, 2025Updated last year
- [AAAI 2025] RhythmMamba☆111Jul 29, 2025Updated last year
- Code implement for FastToG☆92Apr 13, 2025Updated last year
- [AAAI'2025] The official implementation code of SIGMA☆41Oct 14, 2025Updated 10 months ago
- An official implementation of "Re-Attentional Controllable Video Diffusion Editing" in PyTorch. (AAAI 2025)☆27Dec 18, 2024Updated last year
- ☆40Jul 20, 2024Updated 2 years ago
- ☆71Dec 18, 2024Updated last year
- 【AAAI2025】MambaPro: Multi-Modal Object Re-Identification with Mamba Aggregation and Synergistic Prompt☆91May 13, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- AAAI 2025: Hierarchical Consensus Network for Multiview Feature Learning☆18Feb 5, 2025Updated last year
- 【AAAI2025】DeMo: Decoupled Feature-Based Mixture of Experts for Multi-Modal Object Re-Identification☆76Mar 9, 2025Updated last year
- AAAI 2025 (Oral), BrainGuard: Privacy-Preserving Multisubject Image Reconstructions from Brain Activities☆22Dec 1, 2025Updated 9 months ago
- Official implementation of the paper "Attentive Eraser: Unleashing Diffusion Model’s Object Removal Potential via Self-Attention Redirect…☆221Jun 16, 2026Updated 2 months ago
- AAAI 2025: Autonomous LLM-enhanced adversarial attack for text-to-motion☆19Sep 15, 2025Updated 11 months ago
- 【CVPR2025】IDEA: Inverted Text with Cooperative Deformable Aggregation for Multi-modal Object Re-Identification☆52Apr 8, 2025Updated last year
- [NeurIPS 2024] 🕺IMAGPose🕺: A Unified Conditional Framework for Pose-Guided Person Generation. IMAGPose enables versatile pose-guided im…☆169Sep 30, 2025Updated 11 months ago
- [2025] Language-driven Motion Prior Knowledge Learning for Moving Infrared Small Target Detection☆49Jun 17, 2026Updated 2 months ago
- Official code for CustAny: Customizing Anything from A Single Example. Accepted by CVPR2025 (Oral)☆48Apr 10, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repo is the official implementation of "Retrieval-Augmented Dynamic Prompt Tuning for Incomplete Multimodal Learning" accepted by AA…☆67May 26, 2026Updated 3 months ago
- TheaterGen: Character Management with LLM for Consistent Multi-turn Image Generation☆69Sep 26, 2024Updated last year
- ☆27Mar 16, 2026Updated 5 months ago
- [ECCV2024] StoryImager: A Unified and Efficient Framework for Coherent Story Visualization and Completion☆40Jul 5, 2024Updated 2 years ago
- Implementation code:Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models☆192Sep 30, 2025Updated 11 months ago
- [NeurIPS 2024] MuDI: Identity Decoupling for Multi-Subject Personalization of Text-to-Image Models☆96Jan 17, 2025Updated last year
- [AAAI 2025]👔IMAGDressing👔: Interactive Modular Apparel Generation for Virtual Dressing. It enables customizable human image generation …☆1,343Sep 30, 2025Updated 11 months ago
- AugTarget data augmentation for infrared small target detection.☆20May 19, 2023Updated 3 years ago
- [ICCV 2025] MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation☆56Oct 14, 2025Updated 10 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ECCV24] Attention Regulation on T2I Diffusion Models☆19Jul 8, 2024Updated 2 years ago
- [AAAI 2025] Official implementation of the paper "EOV-Seg: Efficient Open-Vocabulary Panoptic Segmentation"☆40Dec 17, 2024Updated last year
- [CVPR 2025] Official PyTorch implementation of StoryGPT-V☆42Jun 14, 2025Updated last year
- [AAAI 2025] GFlow: Recovering 4D World from Monocular Video☆74May 8, 2025Updated last year
- [AAAI2025] Official implementation of the paper "RAP-SR: RestorAtion Prior Enhancement in Diffusion Models for Realistic Image Super-Reso…☆17Mar 22, 2025Updated last year
- [ECCV 2020 Workshop] VIPirios Object Detection Champion☆44Jul 10, 2023Updated 3 years ago
- ☆17Jul 23, 2024Updated 2 years ago