Official code repository of Shuffle-R1
☆26Feb 23, 2026Updated 4 months ago
Alternatives and similar repositories for Shuffle-R1
Users that are interested in Shuffle-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 23] A Simple Vision Transformer for Weakly Semi-supervised 3D Object Detection☆13Apr 12, 2024Updated 2 years ago
- [ECCV 26] Video Streaming Thinking☆114Jun 18, 2026Updated last month
- [CVPR 2026] When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models☆68Apr 11, 2026Updated 3 months ago
- [CVPR 2026] PointTPA: Dynamic Network Parameter Adaptation for 3D Scene Understanding☆33Apr 7, 2026Updated 3 months ago
- HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation☆65May 1, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code release for Dilated-Scale-Aware Category-Attention ConvNet for Multi-Class Object Counting☆22Mar 15, 2023Updated 3 years ago
- Awesome GPT-4 with Applications. This is a collection of resources related to GPT-4, including news, official documents, demo and applica…☆20Mar 15, 2023Updated 3 years ago
- [NeurIPS 2024] A Unified Framework for 3D Scene Understanding☆179Jul 7, 2025Updated last year
- [ECCV 2024] Make Your ViT-based Multi-view 3D Detectors Faster via Token Compression☆53Sep 21, 2024Updated last year
- [ICLR26] ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding☆93Mar 20, 2026Updated 4 months ago
- [NeurIPS 2025] More Than Generation: Unifying Generation and Depth Estimation via Text-to-Image Diffusion Models☆219Oct 31, 2025Updated 8 months ago
- Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models☆266Apr 29, 2026Updated 2 months ago
- [NeurIPS 24] MoE Jetpack: From Dense Checkpoints to Adaptive Mixture of Experts for Vision Tasks☆137Nov 23, 2024Updated last year
- [ECCV 2026] Towards Generalizable Robotic Manipulation in Dynamic Environments☆226Jun 30, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ECCV 2026] Official code of “MindDrive: A Vision-Language-Action Model for Autonomous Driving via Online Reinforcement Learning”☆243Jun 23, 2026Updated 3 weeks ago
- ☆20Jul 21, 2025Updated last year
- [NeurIPS 2025] NoisyRollout: Reinforcing Visual Reasoning with Data Augmentation☆112Sep 18, 2025Updated 10 months ago
- 小样本跨域学习-模式识别大作业☆28May 10, 2024Updated 2 years ago
- [ICCV 2025] A Benchmark for Multi-Step Reasoning in Long Narrative Videos☆28Jun 4, 2026Updated last month
- the official code of DriveMonkey☆45Mar 20, 2026Updated 4 months ago
- Code For Our Work: DVIS-DAQ: Improving Video Segmentation via Dynamic Anchor Queries [ECCV-2024]☆15Jul 11, 2024Updated 2 years ago
- [ICCV 2025] Factorized Learning for Temporally Grounded Video-Language Models☆24Apr 18, 2026Updated 3 months ago
- [AAAI 2026 Oral] Cook and Clean Together: Teaching Embodied Agents for Parallel Task Execution☆363Dec 12, 2025Updated 7 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ECCV 2026] Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding☆421Jun 18, 2026Updated last month
- [ECCV2024] PartGLEE: A Foundation Model for Recognizing and Parsing Any Objects☆64Sep 17, 2024Updated last year
- [ICRA 2026] UniFuture: A 4D Driving World Model for Future Generation and Perception☆162Feb 26, 2026Updated 4 months ago
- [CVPR 2024] Dynamic Adapter Meets Prompt Tuning: Parameter-Efficient Transfer Learning for Point Cloud Analysis☆171Oct 11, 2024Updated last year
- [AAAI 2023 Oral] Language-Assisted 3D Feature Learning for Semantic Scene Understanding☆12Aug 1, 2023Updated 2 years ago
- [ICCV 2025] Implementation of the paper "Q-Frame: Query-aware Frame Selection and Multi-Resolution Adaptation for Video-LLMs"☆81Oct 25, 2025Updated 8 months ago
- ☆24Nov 4, 2025Updated 8 months ago
- Less is Enough: Training-Free Video Diffusion Acceleration via Runtime-Adaptive Caching☆291May 12, 2026Updated 2 months ago
- ☆27Jul 5, 2026Updated 2 weeks ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- This is the official repository of Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities☆42Apr 28, 2026Updated 2 months ago
- ☆21Apr 16, 2025Updated last year
- The official repository for CVPR'26 Paper "APPO: Attention-guided Perception Policy Optimization for Video Reasoning"☆16Mar 19, 2026Updated 4 months ago
- [NeurIPS-2024] The offical Implementation of "Instruction-Guided Visual Masking"☆42Nov 15, 2024Updated last year
- Official implementation of the paper "Instructions are all you need: Self-supervised Reinforcement Learning for Instruction Following"☆40Jan 11, 2026Updated 6 months ago
- This is for ACL 2025 Findings Paper: From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalitiesModels☆102Mar 22, 2026Updated 3 months ago
- [IEEE TPAMI] Parameter-Efficient Fine-Tuning in Spectral Domain for Point Cloud Learning☆391Apr 30, 2026Updated 2 months ago