[ICLR2026] Video-STAR: Reinforcing Open-Vocabulary Action Recognition with Tools
☆208Apr 17, 2026Updated 3 months ago
Alternatives and similar repositories for Video-STAR
Users that are interested in Video-STAR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is the official repository of "LLaTiSA: Towards Difficulty-Stratified Time Series Reasoning from Visual Perception to Semantics".☆78Apr 24, 2026Updated 3 months ago
- ☆61Feb 9, 2026Updated 6 months ago
- [ACM MM 2026] Peak-End-Net: A Peak-End Rule Inspired Framework for Generalizable Video Aesthetic Assessment☆26Updated this week
- ☆60Jun 30, 2026Updated last month
- [CVPR 2026] Elucidating the SNR-t Bias of Diffusion Probabilistic Models☆121Apr 20, 2026Updated 3 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A comprehensive benchmark specifically designed to evaluate the interactive response capabilities of world models in 4D settings.☆107Mar 24, 2026Updated 4 months ago
- [CVPR 2026 Findings] Eevee: Towards Close-up High-resolution Video-based Virtual Try-on☆76Feb 27, 2026Updated 5 months ago
- [SIGGRAPH 2026] MACE-Dance: Motion-Appearance Cascaded Experts for Music-Driven Dance Video Generation☆108May 19, 2026Updated 2 months ago
- [2026 CVPR]Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation☆109Apr 15, 2026Updated 3 months ago
- ☆55Jun 3, 2026Updated 2 months ago
- [CVPR 26] From Scale to Speed: Adaptive Test-Time Scaling for Image Editing☆49Jul 14, 2026Updated 3 weeks ago
- [EMNLP’ 25] Official code for "HS-STaR: Hierarchical Sampling for Self-Taught Reasoners via Difficulty Estimation and Budget Reallocation…☆37Nov 3, 2025Updated 9 months ago
- ☆73Jun 11, 2026Updated 2 months ago
- [ICLR2026] Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models☆145Jan 30, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR 2026] Harder Is Better: Boosting Mathematical Reasoning via Difficulty-Aware GRPO and Multi-Aspect Question Reformulation☆128May 17, 2026Updated 2 months ago
- TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation☆125May 30, 2026Updated 2 months ago
- [AAAI2026] ImagerySearch: Adaptive Test-Time Search for Video Generation Beyond Semantic Dependency Constraints☆56Oct 23, 2025Updated 9 months ago
- [ACL 2026 Findings] Thinking with Map: Reinforced Parallel Map-Augmented Agent for Geolocalization☆177Mar 9, 2026Updated 5 months ago
- [ICLR26] Official implementation of the paper "Urban Socio-Semantic Segmentation with Vision-Language Reasoning"☆175Mar 12, 2026Updated 4 months ago
- [EMNLP25] Official code for "POSITION BIAS MITIGATES POSITION BIAS: Mitigate Position Bias Through Inter-Position Knowledge Distillation…☆38Nov 11, 2025Updated 9 months ago
- [KDD 2026 Oral] MobilityBench: A Scalable Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios☆157Jul 16, 2026Updated 3 weeks ago
- [ICML 2026] The official implementation of paper "Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation…☆91May 25, 2026Updated 2 months ago
- [CVPR 2026] PromptEnhancer is a prompt-rewriting tool, refining prompts into clearer, structured versions for better image generation.☆3,751Jun 10, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- TVM Documentation in Chinese Simplified / TVM 中文文档☆3,895May 20, 2026Updated 2 months ago
- [ICLR2026] AutoDrive-R2: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving☆215May 20, 2026Updated 2 months ago
- LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence https://arxiv.org/abs/2509.03505☆3,929Jun 16, 2026Updated last month
- [ICLR 2026] FASA: FREQUENCY-AWARE SPARSE ATTENTION☆20Mar 1, 2026Updated 5 months ago
- [ECCV 2026] RL3DEdit☆203Jun 30, 2026Updated last month
- [ICLR26] NarrLV: Towards a Comprehensive Narrative-Centric Evaluation for Long Video Generation Models☆112Jul 28, 2025Updated last year
- OpenCode Cover.☆1,624Jun 17, 2026Updated last month
- ☆24Feb 2, 2026Updated 6 months ago
- [ACM MM 25] FingER: Content Aware Fine-grained Evaluation with Reasoning for AI-Generated Videos☆17Jul 17, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Translate PDF, Word, PowerPoint, etc. | zotero翻译插件,微信扫码注册,新用户可免费翻译25万汉字或100万个英文字母。超能文献官网:suppr.wilddata.cn;☆2,011Jun 24, 2026Updated last month
- ☆1,842Feb 14, 2026Updated 5 months ago
- Nexent is a zero-code platform for auto-generating production-grade AI agents using Harness Engineering principles — unified tools, skill…☆5,825Updated this week
- Res-SAM Framework for GPR Underground Hazard Detection☆1,621Jun 15, 2026Updated last month
- Synthetic Data Generation Platform By DataArcTech☆1,772Jun 30, 2026Updated last month
- An agent capable of self-evolving and dynamically hardening security☆2,550Jul 27, 2026Updated 2 weeks ago
- GigaBrain-0: A World Model-Powered Vision-Language-Action Model☆2,560Mar 10, 2026Updated 5 months ago