Self-evolving vision language models from zero data
☆83Mar 14, 2026Updated 6 months ago
Alternatives and similar repositories for MM-Zero
Users that are interested in MM-Zero are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Synthetic Video hallucination and Mitigation☆25Sep 21, 2025Updated 11 months ago
- grpo to train long form QA and instructions with long-form reward model☆17Jul 17, 2025Updated last year
- COS-PLAY: Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Game Play☆31Jul 11, 2026Updated 2 months ago
- ☆20Jul 1, 2026Updated 2 months ago
- Reinforcement Learning of Vision Language Models with Self Visual Perception Reward☆186Mar 14, 2026Updated 6 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The offical repo for "Parallel-Probe: Towards Efficient Parallel Thinking via 2D Probing"☆20Feb 3, 2026Updated 7 months ago
- [ECCV 2026] Official repository of "Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning".☆26Jul 17, 2026Updated 2 months ago
- [CVPR'26] VisPlay: Self-Evolving Vision-Language Models☆77Feb 25, 2026Updated 6 months ago
- The code implementation for TTCS: Test-Time Curriculum Synthesis for Self-Evolving.☆51Apr 22, 2026Updated 4 months ago
- Official implementation for the paper "Video-Based Reward Modeling for Computer-Use Agents"☆17Mar 14, 2026Updated 6 months ago
- This is the codes of "DARE: Aligning LLM Agents with the R Statistical Ecosystem via Distribution-Aware Retrieval"☆15Aug 11, 2026Updated last month
- [ICML'26] VideoGPA is a self-supervised framework that enhances 3D consistency in Video Diffusion Models.☆77Jun 6, 2026Updated 3 months ago
- MemSyco-Bench: Benchmarking Sycophancy in Agent Memory☆20Sep 11, 2026Updated last week
- [CVPR 2026] Official repo for "VideoSSR: Video Self-Supervised Reinforcement Learning"☆48Nov 11, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The released data for paper "Measuring and Improving Chain-of-Thought Reasoning in Vision-Language Models".☆34Sep 16, 2023Updated 3 years ago
- [ICML 2026] Official implementation for paper: Learning Self-Correction in Vision–Language Models via Rollout Augmentation☆16Jun 4, 2026Updated 3 months ago
- A Survey of Self-Evolving Agents | A curated list of resources (surveys, papers, benchmarks, and opensource projects) on Self-Evolving Ag…☆444Sep 12, 2026Updated last week
- OpenClaw-style theorem proving☆29Aug 21, 2026Updated 3 weeks ago
- Build coherent and visually polished multimodal webpages with hierarchical planning, AIGC tools, and iterative reflection.☆18May 17, 2026Updated 4 months ago
- ☆17Mar 10, 2026Updated 6 months ago
- [CVPRW'26] A collection and survey of 3d dataset☆36Jun 4, 2026Updated 3 months ago
- [EMNLP 2026] MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆30Sep 11, 2026Updated last week
- [ICLR2026] codes for R-Zero: Self-Evolving Reasoning LLM from Zero Data (https://www.arxiv.org/pdf/2508.05004)☆850Feb 4, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [COLM 2026] From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement.☆192Aug 12, 2026Updated last month
- [EMNLP '26] Dependency-Aware Structural Retrieval for Massive Agent Skills☆209Aug 30, 2026Updated 2 weeks ago
- Official project page and code repository for WiT, a pixel space diffusion☆17Updated this week
- [CVPR 2026] IOMM: Fast Pre-training of Unified Multimodal Models without Text-Image Pairs☆26Apr 11, 2026Updated 5 months ago
- ☆38Apr 3, 2026Updated 5 months ago
- Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models☆45Jun 14, 2024Updated 2 years ago
- A critical analysis of the Cambrian-S model and VSI-Super benchmarks☆16Nov 20, 2025Updated 10 months ago
- AutoHallusion Codebase (EMNLP 2024)☆22Dec 6, 2024Updated last year
- [ICML 2026] Residual Context Diffusion (RCD): Repurposing discarded signals as structured priors for high-performance reasoning in dLLMs.☆61Jun 28, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SuperDebug,debug如此简单!☆17Jul 19, 2022Updated 4 years ago
- [ICLR 2026] Official repo for "FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting"☆56Oct 9, 2025Updated 11 months ago
- We revisit the Platonic Representation Hypothesis using calibrated representational similarity metrics with statistical guarantees.☆38Jun 24, 2026Updated 2 months ago
- ☆27May 14, 2026Updated 4 months ago
- A unified framework for vision-language environments with Gymnasium-compatible interface☆38Mar 17, 2026Updated 6 months ago
- ☆16Jun 17, 2026Updated 3 months ago
- AcademiClaw: When Students Set Challenges for AI Agents — a bilingual benchmark of 80 university student-sourced academic tasks.☆18Jun 26, 2026Updated 2 months ago