Official implementation of Seeing with You: Perception-Reasoning Co-evolution for Multimodal Reasoning.
☆29Jul 2, 2026Updated 3 months ago
Alternatives and similar repositories for PRCO
Users that are interested in PRCO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of Visco-Attack (EMNLP 2025 Main). An open-source one-click reproduction script is also provided.☆31Apr 11, 2026Updated 5 months ago
- All-in-One Safety Evaluation Framwork☆56Aug 12, 2026Updated last month
- Code implementation for paper "Can Large Language Models Empower Molecular Property Prediction?"☆39Jul 14, 2023Updated 3 years ago
- Official Code for paper "Active Video Perception: Iterative Evidence Seeking for Agentic Long Video Understanding""☆20Jun 2, 2026Updated 4 months ago
- ☆16Jan 12, 2026Updated 8 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆11Oct 25, 2024Updated last year
- Official repo for "PAPO: Perception-Aware Policy Optimization for Multimodal Reasoning"☆163Feb 4, 2026Updated 8 months ago
- ☆29Mar 14, 2026Updated 6 months ago
- Multi-step reasoning MLLM☆26Mar 8, 2026Updated 7 months ago
- Research on "Many-Shot Jailbreaking" in Large Language Models (LLMs). It unveils a novel technique capable of bypassing the safety mechan…☆18Aug 6, 2024Updated 2 years ago
- ☆78Apr 12, 2026Updated 5 months ago
- [NeurIPS 2025] TL;DR: Aligning pretrained unimodal models with the proposed framework using limited paired data yields ~52% gains in cros…☆25Feb 2, 2026Updated 8 months ago
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".☆98Jul 10, 2025Updated last year
- CVPR2025☆23Aug 16, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Vero: An Open RL Recipe for General Visual Reasoning☆149Aug 29, 2026Updated last month
- ☆23Aug 25, 2026Updated last month
- [CVPR 2026] AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition☆45Apr 27, 2026Updated 5 months ago
- ☆34Mar 17, 2026Updated 6 months ago
- ViLoMem: Agentic Learner with Grow-and-Refine Multimodal Semantic Memory☆69Apr 21, 2026Updated 5 months ago
- "Parallel Test-Time Scaling for Latent Reasoning Models"☆24Apr 12, 2026Updated 5 months ago
- Source code for "TeaRAG: A Token-Efficient Agentic Retrieval-Augmented Generation Framework"☆45Jul 2, 2026Updated 3 months ago
- ☆13May 6, 2025Updated last year
- CVPR2025-Multi-party Collaborative Attention Control for Image Customization☆17May 14, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"☆88May 12, 2026Updated 4 months ago
- Official evaluation toolkit of VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis☆16Nov 27, 2025Updated 10 months ago
- [ACL 2026] VGPO: Visually-Guided Policy Optimization for Multimodal Reasoning☆36Apr 14, 2026Updated 5 months ago
- On Policy Distillation Build on top of Verl☆101Sep 3, 2026Updated last month
- ☆41Sep 9, 2025Updated last year
- Code for accepted paper at ICLR 2026☆18Sep 22, 2026Updated 2 weeks ago
- A lifecycle guard skill.☆179Aug 14, 2026Updated last month
- SDPG: Self-Distilled Policy Gradient☆55Jun 15, 2026Updated 3 months ago
- ☆14Feb 26, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Open-source strong baseline for domain generlization re-ID. We will udpate the strong baseline and CFD method~☆10Nov 30, 2021Updated 4 years ago
- The code implementation for TTCS: Test-Time Curriculum Synthesis for Self-Evolving.☆53Apr 22, 2026Updated 5 months ago
- OR-R1: Automating Modeling and Solving of Operations Research Optimization Problem via Test-Time Reinforcement Learning☆19Nov 12, 2025Updated 10 months ago
- VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs☆20Feb 3, 2026Updated 8 months ago
- A collection of the latest research and resources on Fine-Grained Multimodal Perception☆34Jun 4, 2026Updated 4 months ago
- ☆24Dec 21, 2025Updated 9 months ago
- Self-supervised adversarial masking for point clouds☆11Jul 12, 2023Updated 3 years ago