Official code repository of Shuffle-R1
☆26Feb 23, 2026Updated 7 months ago
Alternatives and similar repositories for Shuffle-R1
Users that are interested in Shuffle-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 23] A Simple Vision Transformer for Weakly Semi-supervised 3D Object Detection☆13Apr 12, 2024Updated 2 years ago
- [CVPR 2026] When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models☆70Apr 11, 2026Updated 5 months ago
- [CVPR 2026] PointTPA: Dynamic Network Parameter Adaptation for 3D Scene Understanding☆35Apr 7, 2026Updated 6 months ago
- [IEEE TPAMI] HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation☆72Sep 17, 2026Updated 3 weeks ago
- Code release for Dilated-Scale-Aware Category-Attention ConvNet for Multi-Class Object Counting☆22Mar 15, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Awesome GPT-4 with Applications. This is a collection of resources related to GPT-4, including news, official documents, demo and applica…☆20Mar 15, 2023Updated 3 years ago
- [NeurIPS 2024] A Unified Framework for 3D Scene Understanding☆182Jul 7, 2025Updated last year
- [ECCV 2024] Make Your ViT-based Multi-view 3D Detectors Faster via Token Compression☆53Sep 21, 2024Updated 2 years ago
- [ICLR26] ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding☆87Mar 20, 2026Updated 6 months ago
- ☆12Mar 22, 2025Updated last year
- [NeurIPS 2025] More Than Generation: Unifying Generation and Depth Estimation via Text-to-Image Diffusion Models☆222Oct 31, 2025Updated 11 months ago
- [NeurIPS 24] MoE Jetpack: From Dense Checkpoints to Adaptive Mixture of Experts for Vision Tasks☆139Nov 23, 2024Updated last year
- [ECCV 2026] Towards Generalizable Robotic Manipulation in Dynamic Environments☆238Aug 18, 2026Updated last month
- [ECCV 2026] Official code of “MindDrive: A Vision-Language-Action Model for Autonomous Driving via Online Reinforcement Learning”☆275Jun 23, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICCV 2025] A Benchmark for Multi-Step Reasoning in Long Narrative Videos☆28Jun 4, 2026Updated 4 months ago
- the official code of DriveMonkey☆47Mar 20, 2026Updated 6 months ago
- OpenVLThinker [NeurIPS 2025] & OpenVLThinkerV2 [COLM 2026]☆158May 25, 2026Updated 4 months ago
- Code For Our Work: DVIS-DAQ: Improving Video Segmentation via Dynamic Anchor Queries [ECCV-2024]☆15Jul 11, 2024Updated 2 years ago
- [ICCV 2025] Factorized Learning for Temporally Grounded Video-Language Models☆24Apr 18, 2026Updated 5 months ago
- [AAAI 2026 Oral] Cook and Clean Together: Teaching Embodied Agents for Parallel Task Execution☆366Dec 12, 2025Updated 9 months ago
- [ECCV 2026] Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding☆423Jun 18, 2026Updated 3 months ago
- [AAAI 2023 Oral] Language-Assisted 3D Feature Learning for Semantic Scene Understanding☆12Aug 1, 2023Updated 3 years ago
- [CVPR 2024] Dynamic Adapter Meets Prompt Tuning: Parameter-Efficient Transfer Learning for Point Cloud Analysis☆171Oct 11, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [NeurIPS 2025] NoisyRollout: Reinforcing Visual Reasoning with Data Augmentation☆113Sep 18, 2025Updated last year
- Less is Enough: Training-Free Video Diffusion Acceleration via Runtime-Adaptive Caching☆300May 12, 2026Updated 4 months ago
- [CVPR 2025] A Unified Image-Dense Annotation Generation Model for Underwater Scenes☆62Apr 9, 2025Updated last year
- [ICCV 2025] HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation☆264May 12, 2026Updated 4 months ago
- ☆27Jul 5, 2026Updated 3 months ago
- This is the official repository of Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities☆48Jul 26, 2026Updated 2 months ago
- The official repository for CVPR'26 Paper "APPO: Attention-guided Perception Policy Optimization for Video Reasoning"☆16Mar 19, 2026Updated 6 months ago
- [ICCV 2025] Official code of "ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation"☆667Jun 22, 2026Updated 3 months ago
- Official implementation of the paper "Instructions are all you need: Self-supervised Reinforcement Learning for Instruction Following"☆40Jan 11, 2026Updated 8 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [TMLR 25] SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models☆147Oct 10, 2025Updated 11 months ago
- [ECCV2024] PartGLEE: A Foundation Model for Recognizing and Parsing Any Objects☆66Sep 17, 2024Updated 2 years ago
- Quantized training of Stable Diffusion 3 Medium to significantly reduce memory usage.☆16Jul 10, 2024Updated 2 years ago
- This is for ACL 2025 Findings Paper: From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalitiesModels☆104Mar 22, 2026Updated 6 months ago
- [IEEE TPAMI] Parameter-Efficient Fine-Tuning in Spectral Domain for Point Cloud Learning☆389Apr 30, 2026Updated 5 months ago
- The official code of "PixelWorld: Towards Perceiving Everything as Pixels" [TMLR25]☆15Sep 12, 2025Updated last year
- VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation [TMLR26]☆15Jun 1, 2026Updated 4 months ago