TokenAR: Multiple Subject Generation via Autoregressive Token-level enhancement
☆21Aug 4, 2026Updated last month
Alternatives and similar repositories for TokenAR
Users that are interested in TokenAR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026] Boosting Reasoning in Large Multimodal Models via Activation Replay☆23Aug 25, 2026Updated 3 weeks ago
- ☆23May 26, 2025Updated last year
- Official repository for "Human-MME: A Holistic Evaluation Benchmark for Human-Centric Multimodal Large Language Models"☆23Dec 2, 2025Updated 9 months ago
- Official code repository for Med-CMR : "A Fine-Grained Benchmark Integrating Visual Evidence and Clinical Logic for Medical Complex Multi…☆27Dec 10, 2025Updated 9 months ago
- ☆18Jul 31, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ICML 2026] The official code of FeRA: Frequency–Energy Constrained Routing for Effective Diffusion Adaptation Fine-Tuning☆28Dec 27, 2025Updated 8 months ago
- From Large Angles to Consistent Faces: Identity-Preserving Video Generation via Mixture of Facial Experts☆27Jan 12, 2026Updated 8 months ago
- ☆29Nov 28, 2025Updated 9 months ago
- [ICLR 26] Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual Flow☆46Oct 3, 2025Updated 11 months ago
- ☆93Feb 5, 2026Updated 7 months ago
- ☆185Jun 8, 2026Updated 3 months ago
- [ICML 2026] Transform Trained Transformer for Accelerating Native 4K Video Generation☆41Dec 16, 2025Updated 9 months ago
- CLEAR: Context-Aware Learning with End-to-End Mask-Free Inference for Adaptive Video Subtitle Removal☆25May 25, 2026Updated 3 months ago
- [CVPR2025] Unveil Inversion and Invariance in Flow Transformer for Versatile Image Editing☆25Aug 23, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [CVPR 2026] Soul: Breathe Life into Digital Human for High-fidelity Long-term Multimodal Animation☆64Dec 16, 2025Updated 9 months ago
- [EMNLP 2026 Findings] The official code of Refinement Provenance Inference: Detecting LLM-Refined Training Prompts from Model Behavior☆22Jan 6, 2026Updated 8 months ago
- ☆32Nov 4, 2025Updated 10 months ago
- [ICCV 2025] Official implementation of "Anchor Token Matching: Implicit Structure Locking for Training-free AR Image Editing"☆28Apr 15, 2025Updated last year
- Official code of "Edit Transfer: Learning Image Editing via Vision In-Context Relations"☆89Jun 6, 2025Updated last year
- [CVPR 2025] DreamRelation: Bridging Customization and Relation Generation☆19Dec 17, 2025Updated 9 months ago
- ☆21Apr 17, 2025Updated last year
- ☆42Nov 12, 2025Updated 10 months ago
- Complex-Edit: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark☆30Apr 22, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆31Jan 11, 2026Updated 8 months ago
- Virtual Try-on based on the powerful Flux model☆27Dec 4, 2024Updated last year
- ☆18Mar 14, 2026Updated 6 months ago
- Introduction about AWESOME_ENTROPY+LRM_PAPERS☆34Dec 16, 2025Updated 9 months ago
- RealGeneral (ICCV2025)☆17Jul 16, 2025Updated last year
- [CVPR 2026 Main] MultiBanana: A Challenging Benchmark for Multi-Reference Text-to-Image Generation☆29Aug 4, 2026Updated last month
- Official implementation of "FitDiT: Advancing the Authentic Garment Details for High-fidelity Virtual Try-on"☆630Feb 8, 2025Updated last year
- [ICCV 2025] FreeFlux: Understanding and Exploiting Layer-Specific Roles in RoPE-Based MMDiT for Versatile Image Editing☆77Mar 7, 2026Updated 6 months ago
- ICML2025☆62Aug 28, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Video-CoM: Interactive Video Reasoning via Chain of Manipulations☆23Sep 5, 2026Updated 2 weeks ago
- ☆15Apr 6, 2026Updated 5 months ago
- ☆27Apr 25, 2025Updated last year
- Official repository of the paper InstructBrush: Learning Attention-based Instruction Optimization for Image Editing☆15Apr 14, 2024Updated 2 years ago
- Implementation code of the paper MIGE: A Unified Framework for Multimodal Instruction-Based Image Generation and Editing☆72Jul 13, 2025Updated last year
- OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing☆53Apr 15, 2026Updated 5 months ago
- [NeurIPS 2023] and [ICLR 2024] for robustness certification.☆10Nov 30, 2024Updated last year