Official code repository of Shuffle-R1
☆26Feb 23, 2026Updated 6 months ago
Alternatives and similar repositories for Shuffle-R1
Users that are interested in Shuffle-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 23] A Simple Vision Transformer for Weakly Semi-supervised 3D Object Detection☆13Apr 12, 2024Updated 2 years ago
- [ECCV 26] Video Streaming Thinking☆124Jul 28, 2026Updated last month
- [CVPR 2026] When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models☆70Apr 11, 2026Updated 5 months ago
- [CVPR 2026] PointTPA: Dynamic Network Parameter Adaptation for 3D Scene Understanding☆35Apr 7, 2026Updated 5 months ago
- [IEEE TPAMI] HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation☆71Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code release for Dilated-Scale-Aware Category-Attention ConvNet for Multi-Class Object Counting☆22Mar 15, 2023Updated 3 years ago
- [NeurIPS 2024] A Unified Framework for 3D Scene Understanding☆180Jul 7, 2025Updated last year
- [ECCV 2024] Make Your ViT-based Multi-view 3D Detectors Faster via Token Compression☆53Sep 21, 2024Updated last year
- [ICLR26] ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding☆87Mar 20, 2026Updated 5 months ago
- ☆12Mar 22, 2025Updated last year
- [NeurIPS 2025] More Than Generation: Unifying Generation and Depth Estimation via Text-to-Image Diffusion Models☆222Oct 31, 2025Updated 10 months ago
- [NeurIPS 24] MoE Jetpack: From Dense Checkpoints to Adaptive Mixture of Experts for Vision Tasks☆138Nov 23, 2024Updated last year
- [ECCV 2026] Towards Generalizable Robotic Manipulation in Dynamic Environments☆230Aug 18, 2026Updated last month
- [ECCV 2026] Official code of “MindDrive: A Vision-Language-Action Model for Autonomous Driving via Online Reinforcement Learning”☆267Jun 23, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆20Jul 21, 2025Updated last year
- [ICCV 2025] A Benchmark for Multi-Step Reasoning in Long Narrative Videos☆28Jun 4, 2026Updated 3 months ago
- the official code of DriveMonkey☆47Mar 20, 2026Updated 6 months ago
- OpenVLThinker [NeurIPS 2025] & OpenVLThinkerV2 [COLM 2026]☆156May 25, 2026Updated 3 months ago
- Code For Our Work: DVIS-DAQ: Improving Video Segmentation via Dynamic Anchor Queries [ECCV-2024]☆15Jul 11, 2024Updated 2 years ago
- [ICCV 2025] Factorized Learning for Temporally Grounded Video-Language Models☆24Apr 18, 2026Updated 5 months ago
- [AAAI 2026 Oral] Cook and Clean Together: Teaching Embodied Agents for Parallel Task Execution☆365Dec 12, 2025Updated 9 months ago
- [ECCV 2026] Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding☆421Jun 18, 2026Updated 3 months ago
- [ICRA 2026] UniFuture: A 4D Driving World Model for Future Generation and Perception☆165Feb 26, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [AAAI 2023 Oral] Language-Assisted 3D Feature Learning for Semantic Scene Understanding☆12Aug 1, 2023Updated 3 years ago
- [NeurIPS 2025] NoisyRollout: Reinforcing Visual Reasoning with Data Augmentation☆112Sep 18, 2025Updated last year
- ☆24Nov 4, 2025Updated 10 months ago
- Less is Enough: Training-Free Video Diffusion Acceleration via Runtime-Adaptive Caching☆299May 12, 2026Updated 4 months ago
- [ICCV 2025] HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation☆264May 12, 2026Updated 4 months ago
- [CVPR 2025] A Unified Image-Dense Annotation Generation Model for Underwater Scenes☆61Apr 9, 2025Updated last year
- ☆27Jul 5, 2026Updated 2 months ago
- This is the official repository of Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities☆48Jul 26, 2026Updated last month
- The official repository for CVPR'26 Paper "APPO: Attention-guided Perception Policy Optimization for Video Reasoning"☆16Mar 19, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICCV 2025] Official code of "ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation"☆664Jun 22, 2026Updated 2 months ago
- [TMLR 25] SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models☆147Oct 10, 2025Updated 11 months ago
- [ECCV2024] PartGLEE: A Foundation Model for Recognizing and Parsing Any Objects☆65Sep 17, 2024Updated 2 years ago
- This is for ACL 2025 Findings Paper: From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalitiesModels☆103Mar 22, 2026Updated 5 months ago
- [IEEE TPAMI] Parameter-Efficient Fine-Tuning in Spectral Domain for Point Cloud Learning☆389Apr 30, 2026Updated 4 months ago
- The official code of "PixelWorld: Towards Perceiving Everything as Pixels" [TMLR25]☆15Sep 12, 2025Updated last year
- VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation [TMLR26]☆15Jun 1, 2026Updated 3 months ago