π΅ Code for our EMNLP 2025 Main paper: "FlashAdventure: A Benchmark for GUI Agents Solving Full Story Arcs in Diverse Adventure Games"
β27Apr 26, 2026Updated 2 months ago
Alternatives and similar repositories for FlashAdventure
Users that are interested in FlashAdventure are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β153Jun 17, 2025Updated last year
- Official Repository for our CVPR2024 paper: ESR-NeRF: Emissive Source Reconstruction Using LDR Multi-view Imagesβ15Jun 13, 2024Updated 2 years ago
- π§π» Code and benchmark for our Findings of ACL 2024 paper - "TimeChara: Evaluating Point-in-Time Character Hallucination of Role-Playingβ¦β21Dec 20, 2024Updated last year
- Official code for ICML 2024 paper "Learning to Continually Learn with the Bayesian Principle"β21May 27, 2024Updated 2 years ago
- A Survey on Large Foundation Models as Game Players - Datasets, Models, Harness and Benchmarksβ37May 13, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- π μμΈλ μ»΄ν¨ν°κ³΅νλΆ (컴곡) νμ λ Όλ¬Έ ν νλ¦Ώ | Thesis template for SNU CSEβ20Jan 5, 2026Updated 6 months ago
- Official implementation of "ViSAGe: Video-to-Spatial AUdio Generation" (ICLR 2025)β47Sep 10, 2025Updated 10 months ago
- [ICML 2025] Adaptive Self-improvement LLM Agentic System for ML Library Developmentβ17Jan 6, 2026Updated 6 months ago
- Official repository of EP2P-Loc: End-to-End 3D Point to 2D Pixel Localization for Large-Scale Visual Localization (ICCV 2023)β63Oct 16, 2023Updated 2 years ago
- An Ultra-Long Output Reinforcement Learning Approachβ23Jul 31, 2025Updated 11 months ago
- Official PyTorch implementation of "Efficient Latency-Aware CNN Depth Compression via Two-Stage Dynamic Programming" (ICML'23)β13Apr 13, 2026Updated 3 months ago
- Benchmark environment for evaluating vision-language models (VLMs) on popular video games!β364May 30, 2025Updated last year
- The official implemention of "Depth-Breadth Synergy in RLVR: Unlocking LLM Reasoning Gains with Adaptive Exploration" (ICML 2026)β24Feb 4, 2026Updated 5 months ago
- [ICLR 2024] DMBP: Diffusion Model-Based Predictor for Robust Offline Reinforcement Learning against State Observations Perturbations.β17May 24, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ACL '26 Findings] V-MAGE: A Game Evaluation Framework for Assessing Visual-Centric Capabilities in MLLMsβ27Apr 28, 2026Updated 2 months ago
- β60Oct 18, 2024Updated last year
- A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Modelsβ30Nov 25, 2024Updated last year
- This is the repository for paper EscapeBench: Pushing Language Models to Think Outside the Boxβ18Dec 19, 2024Updated last year
- [COLM 2026] An efficient 3D sampling method for long-CoT LLM.β16May 25, 2025Updated last year
- A secure, zero-touch API provisioning protocol driven by contracts and powered by AIβ13May 4, 2025Updated last year
- β24Dec 30, 2024Updated last year
- β11Jun 3, 2023Updated 3 years ago
- The code and data for "Summary-Oriented Vision Modeling for Multimodal Abstractive Summarization"β11May 16, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official code and dataset for our NAACL 2024 paper: DialogCC: An Automated Pipeline for Creating High-Quality Multi-modal Dialogue Dataseβ¦β13Jun 24, 2024Updated 2 years ago
- ACAV100M: Automatic Curation of Large-Scale Datasets for Audio-Visual Video Representation Learning. In ICCV, 2021.β64Nov 18, 2021Updated 4 years ago
- [CVPR 2024 Highlight] ImageNet-Dβ47Updated this week
- [ACM MM '24 Poster] Official repository of paper titled "Towards Robustness Prompt Tuning with Fully Test-Time Adaptation for CLIPβs Zeroβ¦β10Aug 6, 2024Updated last year
- [ICCV 25] Official repository of "Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dialβ¦β31Apr 1, 2026Updated 3 months ago
- β130Oct 3, 2025Updated 9 months ago
- Code for paper OpenWebRL: Online Multi-Turn Reinforcement Learning for Visual Web Agentsβ37Jul 9, 2026Updated 2 weeks ago
- β76Jun 10, 2025Updated last year
- β16Jan 4, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β12May 17, 2022Updated 4 years ago
- [ICML 2026] GameVerse: Can Vision-Language Models Learn from Video-based Reflection?β51Jul 13, 2026Updated last week
- β11Oct 16, 2023Updated 2 years ago
- Uncertainty-Guided Pseudo-Labelling with Model Averagingβ11Mar 17, 2026Updated 4 months ago
- The source code of ExFunTubeβ10Aug 8, 2025Updated 11 months ago
- Official Implementation for "In-Context Reinforcement Learning from Noise Distillation"β35Sep 18, 2024Updated last year
- The code for paper "EPO: Entropy-regularized Policy Optimization for LLM Agents Reinforcement Learning"β40Jul 13, 2026Updated last week