Self-evolving vision language models from zero data
☆84Mar 14, 2026Updated 6 months ago
Alternatives and similar repositories for MM-Zero
Users that are interested in MM-Zero are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Synthetic Video hallucination and Mitigation☆25Sep 21, 2025Updated last year
- grpo to train long form QA and instructions with long-form reward model☆17Jul 17, 2025Updated last year
- COS-PLAY: Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Game Play☆31Jul 11, 2026Updated 2 months ago
- ☆20Jul 1, 2026Updated 3 months ago
- Reinforcement Learning of Vision Language Models with Self Visual Perception Reward☆185Mar 14, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The offical repo for "Parallel-Probe: Towards Efficient Parallel Thinking via 2D Probing"☆21Feb 3, 2026Updated 8 months ago
- [ECCV 2026] Official repository of "Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning".☆30Updated this week
- [CVPR'26] VisPlay: Self-Evolving Vision-Language Models☆79Feb 25, 2026Updated 7 months ago
- The code implementation for TTCS: Test-Time Curriculum Synthesis for Self-Evolving.☆53Apr 22, 2026Updated 5 months ago
- Video Content Customization Using First Frame☆194Mar 17, 2026Updated 6 months ago
- Official implementation for the paper "Video-Based Reward Modeling for Computer-Use Agents"☆17Mar 14, 2026Updated 6 months ago
- This is the codes of "DARE: Aligning LLM Agents with the R Statistical Ecosystem via Distribution-Aware Retrieval"☆15Aug 11, 2026Updated last month
- [ICML'26] VideoGPA is a self-supervised framework that enhances 3D consistency in Video Diffusion Models.☆79Updated this week
- [CVPR 2026] Official repo for "VideoSSR: Video Self-Supervised Reinforcement Learning"☆48Nov 11, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The released data for paper "Measuring and Improving Chain-of-Thought Reasoning in Vision-Language Models".☆34Sep 16, 2023Updated 3 years ago
- [ICML 2026] Official implementation for paper: Learning Self-Correction in Vision–Language Models via Rollout Augmentation☆16Jun 4, 2026Updated 4 months ago
- A Survey of Self-Evolving Agents | A curated list of resources (surveys, papers, benchmarks, and opensource projects) on Self-Evolving Ag…☆469Updated this week
- A most Frontend Collection and survey of vision-language model papers, and models GitHub repository. Continuous updates.☆723Sep 14, 2026Updated 3 weeks ago
- OpenClaw-style theorem proving☆30Aug 21, 2026Updated last month
- Build coherent and visually polished multimodal webpages with hierarchical planning, AIGC tools, and iterative reflection.☆18May 17, 2026Updated 4 months ago
- ☆17Mar 10, 2026Updated 7 months ago
- [EMNLP 2026] MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆30Updated this week
- [ICLR2026] codes for R-Zero: Self-Evolving Reasoning LLM from Zero Data (https://www.arxiv.org/pdf/2508.05004)☆857Feb 4, 2026Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [COLM 2026] From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement.☆192Aug 12, 2026Updated last month
- Official project page and code repository for WiT, a pixel space diffusion☆17Sep 14, 2026Updated 3 weeks ago
- [CVPR 2026] IOMM: Fast Pre-training of Unified Multimodal Models without Text-Image Pairs☆26Apr 11, 2026Updated 5 months ago
- Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models☆45Jun 14, 2024Updated 2 years ago
- A critical analysis of the Cambrian-S model and VSI-Super benchmarks☆16Nov 20, 2025Updated 10 months ago
- AutoHallusion Codebase (EMNLP 2024)☆22Dec 6, 2024Updated last year
- [ICML 2026] Residual Context Diffusion (RCD): Repurposing discarded signals as structured priors for high-performance reasoning in dLLMs.☆62Jun 28, 2026Updated 3 months ago
- SuperDebug,debug如此简单!☆17Jul 19, 2022Updated 4 years ago
- [ICLR 2026] Official repo for "FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting"☆56Oct 9, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code and Data for "FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented Generation" (ACL25)☆39Oct 26, 2025Updated 11 months ago
- [ICML 2026] We revisit the Platonic Representation Hypothesis using calibrated representational similarity metrics with statistical guara…☆40Jun 24, 2026Updated 3 months ago
- ☆30May 14, 2026Updated 4 months ago
- A unified framework for vision-language environments with Gymnasium-compatible interface☆38Mar 17, 2026Updated 6 months ago
- ☆16Jun 17, 2026Updated 3 months ago
- AcademiClaw: When Students Set Challenges for AI Agents — a bilingual benchmark of 80 university student-sourced academic tasks.☆20Jun 26, 2026Updated 3 months ago
- ☆24Jun 23, 2026Updated 3 months ago