Code for the paper "AutoPresent: Designing Structured Visuals From Scratch" (CVPR 2025)
β181May 26, 2025Updated last year
Alternatives and similar repositories for AutoPresent
Users that are interested in AutoPresent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The code in "SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design"β48Oct 20, 2025Updated 11 months ago
- π₯ [ICML 2026] Official implementation of "Are LRMs Interruptible?"β20Jun 18, 2026Updated 3 months ago
- Echo: "Constantly Improving Image Models Need Constantly Improving Benchmarks" (ICLR 2026)β20Sep 22, 2026Updated 2 weeks ago
- [AAAI 2026] SlideTailor: Personalized Presentation Slide Generation for Scientific Papersβ60Apr 18, 2026Updated 5 months ago
- Recursive Visual Programming (ECCV 2024)β19Nov 20, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official Repository of VisGym: Diverse, Customizable, Scalable Environments for Multimodal Agentsβ117May 3, 2026Updated 5 months ago
- [ICLR 2026] P2P: Automated Paper-to-Poster Generation and Fine-Grained Benchmarkβ56Jun 6, 2025Updated last year
- β24Aug 21, 2026Updated last month
- π₯ [ICLR 2025] Official PyTorch Model "Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark"β26Feb 9, 2025Updated last year
- Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMsβ71Mar 22, 2026Updated 6 months ago
- π₯ [NeurIPS 2025] Official implementation of "Generate, but Verify: Reducing Visual Hallucination in Vision-Language Models with Retrospeβ¦β59Jan 22, 2026Updated 8 months ago
- TPDiff: Temporal Pyramid Video Diffusion Modelβ25Mar 13, 2025Updated last year
- [EMNLP 2025] DiagramEval: Evaluating LLM-Generated Diagrams via Graphsβ17Nov 1, 2025Updated 11 months ago
- [ICCV 2025] Preacher: Paper-to-Video Agentic Systemβ52Sep 1, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- π§© Official code repository for βPuzzled by Puzzles: When Vision-Language Models Canβt Take a Hint.ββ15Sep 22, 2025Updated last year
- β14Mar 18, 2025Updated last year
- Benchmark for Agentic Powerpoint Editing Tasksβ29Sep 25, 2026Updated 2 weeks ago
- [NeurIPS 2024] The official implementation of "Image Copy Detection for Diffusion Models"β18Oct 1, 2024Updated 2 years ago
- β21May 19, 2025Updated last year
- The implementation for ThreadWeaver Adaptive Threading for Efficient Parallel Reasoning in Language Modelsβ67Apr 8, 2026Updated 6 months ago
- [TMLR 2026] Multimodal Large Language Models for Code Generation under Multimodal Scenariosβ282Updated this week
- [ECCV2024] Fast Sprite Decomposition from Animated Graphicsβ32Sep 26, 2024Updated 2 years ago
- Collection of Aesthetics Assessment Papers for Graphic Designs.β47Aug 11, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- (CVPR 2025) Code of "Chat2SVG: Vector Graphics Generation with Large Language Models and Image Diffusion Models"β254Apr 2, 2025Updated last year
- β13Aug 14, 2022Updated 4 years ago
- β72Apr 13, 2026Updated 5 months ago
- [CVPR 2025] PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Modelsβ54Jun 12, 2025Updated last year
- [ICCV 2025] Diffusion Curriculum (DisCL)β19Sep 26, 2025Updated last year
- This repository is for the paper "Is BERT Blind? Exploring the Effect of Vision-and-Language Pretraining on Visual Language Understandingβ¦β21Nov 2, 2023Updated 2 years ago
- [ECCV 2024 Oral] The official implementation of paper: COHO: Context-Sensitive City-Scale Hierarchical Urban Layout Generationβ14Aug 13, 2024Updated 2 years ago
- β22Feb 10, 2025Updated last year
- [ICLR 2025] Aligning Generative Denoising with Discriminative Objectives Unleashes Diffusion for Visual Perceptionβ16Jul 4, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- β31Jan 7, 2024Updated 2 years ago
- β30Feb 27, 2026Updated 7 months ago
- We introduce new approach, Token Reduction using CLIP Metric (TRIM), aimed at improving the efficiency of MLLMs without sacrificing theirβ¦β23Jan 11, 2026Updated 8 months ago
- [NeurIPS2024] Official code for (IMA) Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputsβ23Oct 15, 2024Updated last year
- [ECCV 2026] Official repository of "Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning".β30Updated this week
- β16Jun 14, 2024Updated 2 years ago
- LangCode - Improving alignment and reasoning of large language models (LLMs) with natural language embedded program (NLEP).β51Sep 22, 2023Updated 3 years ago