Self-evolving vision language models from zero data
☆81Mar 14, 2026Updated 5 months ago
Alternatives and similar repositories for MM-Zero
Users that are interested in MM-Zero are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Synthetic Video hallucination and Mitigation☆25Sep 21, 2025Updated 11 months ago
- An easy python package to run quick basic QA evaluations. This package includes standardized QA evaluation metrics and semantic evaluatio…☆63Jul 18, 2025Updated last year
- grpo to train long form QA and instructions with long-form reward model☆17Jul 17, 2025Updated last year
- COS-PLAY: Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Game Play☆30Jul 11, 2026Updated last month
- ☆19Jul 1, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Reinforcement Learning of Vision Language Models with Self Visual Perception Reward☆183Mar 14, 2026Updated 5 months ago
- The offical repo for "Parallel-Probe: Towards Efficient Parallel Thinking via 2D Probing"☆19Feb 3, 2026Updated 6 months ago
- [ECCV 2026] Official repository of "Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning".☆25Jul 17, 2026Updated last month
- [CVPR'26] VisPlay: Self-Evolving Vision-Language Models☆75Feb 25, 2026Updated 6 months ago
- Video Content Customization Using First Frame☆194Mar 17, 2026Updated 5 months ago
- Official implementation for the paper "Video-Based Reward Modeling for Computer-Use Agents"☆17Mar 14, 2026Updated 5 months ago
- This is the codes of "DARE: Aligning LLM Agents with the R Statistical Ecosystem via Distribution-Aware Retrieval"☆15Aug 11, 2026Updated 2 weeks ago
- [ICML'26] VideoGPA is a self-supervised framework that enhances 3D consistency in Video Diffusion Models.☆72Jun 6, 2026Updated 2 months ago
- MemSyco-Bench: Benchmarking Sycophancy in Agent Memory☆17Jul 30, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR 2026] Official repo for "VideoSSR: Video Self-Supervised Reinforcement Learning"☆48Nov 11, 2025Updated 9 months ago
- The released data for paper "Measuring and Improving Chain-of-Thought Reasoning in Vision-Language Models".☆34Sep 16, 2023Updated 2 years ago
- [ICML 2026] Official implementation for paper: Learning Self-Correction in Vision–Language Models via Rollout Augmentation☆16Jun 4, 2026Updated 2 months ago
- A Survey of Self-Evolving Agents | A curated list of resources (surveys, papers, benchmarks, and opensource projects) on Self-Evolving Ag…☆404Aug 23, 2026Updated last week
- A most Frontend Collection and survey of vision-language model papers, and models GitHub repository. Continuous updates.☆704Aug 22, 2026Updated last week
- ☆31Apr 11, 2026Updated 4 months ago
- OpenClaw-style theorem proving☆29Aug 21, 2026Updated last week
- Build coherent and visually polished multimodal webpages with hierarchical planning, AIGC tools, and iterative reflection.☆17May 17, 2026Updated 3 months ago
- ☆17Mar 10, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP 2026] MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆30May 23, 2026Updated 3 months ago
- [ICLR2026] codes for R-Zero: Self-Evolving Reasoning LLM from Zero Data (https://www.arxiv.org/pdf/2508.05004)☆842Feb 4, 2026Updated 6 months ago
- [COLM 2026] From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement.☆191Aug 12, 2026Updated 2 weeks ago
- [EMNLP '26] Dependency-Aware Structural Retrieval for Massive Agent Skills☆202Aug 21, 2026Updated last week
- Official project page and code repository for WiT, a pixel space diffusion☆17May 31, 2026Updated 3 months ago
- [CVPR 2026] IOMM: Fast Pre-training of Unified Multimodal Models without Text-Image Pairs☆26Apr 11, 2026Updated 4 months ago
- ☆36Apr 3, 2026Updated 4 months ago
- Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models☆45Jun 14, 2024Updated 2 years ago
- A critical analysis of the Cambrian-S model and VSI-Super benchmarks☆16Nov 20, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AutoHallusion Codebase (EMNLP 2024)☆22Dec 6, 2024Updated last year
- [ICML 2026] Residual Context Diffusion (RCD): Repurposing discarded signals as structured priors for high-performance reasoning in dLLMs.☆60Jun 28, 2026Updated 2 months ago
- [ICLR 2026] Official repo for "FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting"☆56Oct 9, 2025Updated 10 months ago
- Code and Data for "FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented Generation" (ACL25)☆39Oct 26, 2025Updated 10 months ago
- We revisit the Platonic Representation Hypothesis using calibrated representational similarity metrics with statistical guarantees.☆38Jun 24, 2026Updated 2 months ago
- ☆26May 14, 2026Updated 3 months ago
- A unified framework for vision-language environments with Gymnasium-compatible interface☆38Mar 17, 2026Updated 5 months ago