Benchmark for Agentic Powerpoint Editing Tasks
☆22Jul 6, 2026Updated 3 weeks ago
Alternatives and similar repositories for PPTArena
Users that are interested in PPTArena are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2025] Aligning Generative Denoising with Discriminative Objectives Unleashes Diffusion for Visual Perception☆15Jul 4, 2025Updated last year
- ☆25Mar 30, 2025Updated last year
- [CVPR 2026 Main] MultiBanana: A Challenging Benchmark for Multi-Reference Text-to-Image Generation☆29Updated this week
- [ECCV-24] This is the official implementation of the paper "SEGIC: Unleashing the Emergent Correspondence for In-Context Segmentation".☆27Oct 13, 2024Updated last year
- Official implementation for the paper "Video-Based Reward Modeling for Computer-Use Agents"☆17Mar 14, 2026Updated 4 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- An official implementation of FlashI2V.☆33Nov 16, 2025Updated 8 months ago
- ☆34Dec 29, 2025Updated 7 months ago
- ☆19Jul 28, 2025Updated last year
- [IROS 2023] DualCross: Cross-Modality Cross-Domain Adaptation for Monocular BEV Perception☆32Nov 28, 2023Updated 2 years ago
- Official repository of the paper InstructBrush: Learning Attention-based Instruction Optimization for Image Editing☆16Apr 14, 2024Updated 2 years ago
- ☆15Jun 9, 2025Updated last year
- Code for "Skill-based Chain-of-Thoughts for Domain-Adaptive Video Reasoning [EMNLP 2025 Findings]"☆18Aug 27, 2025Updated 11 months ago
- Developer project for getting basic API integrations working in under 5 minutes☆11May 22, 2026Updated 2 months ago
- Code for LaMPP: Language Models as Probabilistic Priors for Perception and Action☆37Apr 3, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2024] SHAP-EDITOR: Instruction-guided Latent 3D Editing in Seconds☆39Jul 19, 2025Updated last year
- Official Implementation for "SiLVR : A Simple Language-based Video Reasoning Framework"☆19Jan 18, 2026Updated 6 months ago
- A reinforcement learning framework with verifiable aesthetic rewards for improving aesthetic slide generation capabilities in LLM agents.…☆30May 19, 2026Updated 2 months ago
- ☆13Jun 10, 2025Updated last year
- MR. Video: MapReduce is the Principle for Long Video Understanding☆31Jun 18, 2026Updated last month
- [CVPR2026] Official implementation of "FreqEdit: Preserving High-Frequency Features for Robust Multi-Turn Image Editing"☆15Mar 31, 2026Updated 4 months ago
- [CVPR 2025] GPS as a Control Signal for Image Generation☆25Mar 18, 2025Updated last year
- Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Models☆19Jan 21, 2026Updated 6 months ago
- 这里将paddle中的ocr等模型转为onnx格式,并利用java版深度框架djl加载这些onnx模型进行推理预测尝试。☆14Nov 15, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Jul 5, 2024Updated 2 years ago
- ☆17Mar 19, 2026Updated 4 months ago
- Code for paper "OmniMotion-X: Versatile Multimodal Whole-Body Motion Generation"☆31Oct 23, 2025Updated 9 months ago
- VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation☆20Jun 2, 2025Updated last year
- NeurIPS 2025 Poster☆26Feb 4, 2025Updated last year
- Video-CoM: Interactive Video Reasoning via Chain of Manipulations☆22Jun 17, 2026Updated last month
- ☆25May 12, 2026Updated 2 months ago
- This project is the official implementation of 'DreamOmni3: Scribble-based Editing and Generation''☆40Dec 30, 2025Updated 7 months ago
- CLEAR: Context-Aware Learning with End-to-End Mask-Free Inference for Adaptive Video Subtitle Removal☆21May 25, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- (Siggraph Asia 2023) Project Page of "HyperDreamer: Hyper-Realistic 3D Content Generation and Editing from a Single Image"☆10Dec 9, 2023Updated 2 years ago
- ULMEvalKit: One-Stop Eval ToolKit for Image Generation☆56Dec 17, 2025Updated 7 months ago
- The official repo for "VisualWebInstruct: Scaling up Multimodal Instruction Data through Web Search" [EMNLP25]☆39Feb 1, 2026Updated 6 months ago
- ☆19Sep 19, 2024Updated last year
- TokenAR: Multiple Subject Generation via Autoregressive Token-level enhancement☆22Updated this week
- A instruction data generation system for multimodal language models.☆37Jan 31, 2025Updated last year
- Learning from Next-Frame Prediction: Autoregressive Video Modeling Encodes Effective Representations☆22Dec 24, 2025Updated 7 months ago