☆18Jul 31, 2025Updated last year
Alternatives and similar repositories for MACT
Users that are interested in MACT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TokenAR: Multiple Subject Generation via Autoregressive Token-level enhancement☆22Aug 4, 2026Updated 2 months ago
- Official repository for "Human-MME: A Holistic Evaluation Benchmark for Human-Centric Multimodal Large Language Models"☆24Dec 2, 2025Updated 10 months ago
- ☆23May 26, 2025Updated last year
- [CVPR 2026] Boosting Reasoning in Large Multimodal Models via Activation Replay☆26Aug 25, 2026Updated last month
- Official code repository for Med-CMR : "A Fine-Grained Benchmark Integrating Visual Evidence and Clinical Logic for Medical Complex Multi…☆28Dec 10, 2025Updated 9 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR 26] Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual Flow☆47Oct 3, 2025Updated last year
- From Large Angles to Consistent Faces: Identity-Preserving Video Generation via Mixture of Facial Experts☆27Jan 12, 2026Updated 8 months ago
- ☆29Nov 28, 2025Updated 10 months ago
- ☆97Feb 5, 2026Updated 8 months ago
- ☆187Jun 8, 2026Updated 4 months ago
- [CVPR 2026] Soul: Breathe Life into Digital Human for High-fidelity Long-term Multimodal Animation☆64Dec 16, 2025Updated 9 months ago
- [EMNLP 2026 Findings] The official code of Refinement Provenance Inference: Detecting LLM-Refined Training Prompts from Model Behavior☆23Jan 6, 2026Updated 9 months ago
- ☆15Apr 6, 2026Updated 6 months ago
- ☆18Jul 14, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- 4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding☆80May 26, 2026Updated 4 months ago
- ☆32Jan 11, 2026Updated 8 months ago
- [CVPR2025] Unveil Inversion and Invariance in Flow Transformer for Versatile Image Editing☆25Aug 23, 2025Updated last year
- Modality Gap Theory☆76May 16, 2026Updated 4 months ago
- ☆18Jun 3, 2025Updated last year
- [NeurIPS 2023] and [ICLR 2024] for robustness certification.☆10Nov 30, 2024Updated last year
- ☆42Nov 12, 2025Updated 10 months ago
- [EMNLP'26 Findings] OPD-Evolver☆45Jun 17, 2026Updated 3 months ago
- Memory-optimized training scripts for video models based on Diffusers☆17Jan 3, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆10Jan 6, 2025Updated last year
- The evaluation tool (Matlab version) for saliency maps.☆10Mar 18, 2022Updated 4 years ago
- Code Implementation for AutoAttend: Automated Attention Representation Search☆11Jul 26, 2021Updated 5 years ago
- CurriculumLoc for Visual Geo-localization☆16Nov 23, 2023Updated 2 years ago
- VS-Bench: Evaluating VLMs for Strategic Reasoning and Decision-Making in Multi-Agent Environments☆26Sep 30, 2025Updated last year
- ☆17Jan 14, 2026Updated 8 months ago
- ☆21Apr 17, 2025Updated last year
- [CVPR 2025] DreamRelation: Bridging Customization and Relation Generation☆18Dec 17, 2025Updated 9 months ago
- ACM MM Workshop on UAVs in Multimedia: Capturing the World from a New Perspective (UAVM 2023)☆13Jul 4, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Our solution in competition NTIRE 2023 Bokeh Effect Transformation: https://codalab.lisn.upsaclay.fr/competitions/10229☆15Jun 18, 2023Updated 3 years ago
- UAVM @ ACM MM2023 Workshop on UAVs in Multimedia: Capturing the World from a New Perspective☆17Apr 30, 2025Updated last year
- Target-Grounded Graph-Aware Transformer for Aerial Vision-and-Dialog Navigation, AVDN Challenge, ICCV CLVL 2023.☆21Jan 2, 2024Updated 2 years ago
- Unofficial implementation of Layer Diffuse in diffusers☆28Apr 3, 2024Updated 2 years ago
- ☆11Aug 27, 2024Updated 2 years ago
- Self-Supervised Learning for Fine-Grained Image Categorization☆26Dec 18, 2022Updated 3 years ago
- [CVPR 2025] DV-Matcher: Deformation-based Non-Rigid Point Cloud Matching Guided by Pre-trained Visual Features☆28Sep 5, 2025Updated last year