Reproduction code for paper "MineExplorer: Evaluating Open-World Exploration of MLLM Agents in Minecraft"
☆21Jun 12, 2026Updated 2 months ago
Alternatives and similar repositories for MineExplorer
Users that are interested in MineExplorer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reproduction Code for Paper "Investigating Multi-Hop Factual Shortcuts in Knowledge Editing of Large Language Models"☆14Jun 1, 2024Updated 2 years ago
- The official repository of paper "Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark"☆19Jun 20, 2025Updated last year
- [ACM MM 2026] MemGUI-Bench: Benchmarking Memory of Mobile GUI Agents in Dynamic Environments☆48Updated this week
- [ACL'26 Oral] AgentOCR is a token-efficient framework that compresses multi-turn agent history by rendering it into images and adopting R…☆47Mar 1, 2026Updated 5 months ago
- This is the official repository of the paper "Atomic-to-Compositional Generalization for Mobile Agents with A New Benchmark and Schedulin…☆15Jul 27, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning.☆25Oct 7, 2025Updated 10 months ago
- ☆17Feb 26, 2024Updated 2 years ago
- [ACL 2026] LearnAct: Few-Shot Mobile GUI Agent with a Unified Demonstration Benchmark☆49Updated this week
- Collections of RLxLM experiments using minimal codes☆14Feb 17, 2025Updated last year
- NeurIPS 2025☆16Feb 4, 2026Updated 6 months ago
- inductive reasoning benchmark with subregular hierarchy for string-to-string transformation☆20Jun 27, 2025Updated last year
- Awesome Audio-Visual Intelligence, Survey of Audio-Visual Intelligence☆86May 8, 2026Updated 3 months ago
- ☆36Apr 16, 2026Updated 4 months ago
- ☆13Jul 14, 2024Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Official code for Attention-driven GUI Grounding, AAAI2025☆16Dec 17, 2024Updated last year
- Implementation of SLIM, a framework of dynamics skill lifecycle management for agentic reinforcement learning☆22May 12, 2026Updated 3 months ago
- [ICLR'25 Oral] MMIE: Massive Multimodal Interleaved Comprehension Benchmark for Large Vision-Language Models☆35Nov 3, 2024Updated last year
- [ICLR 26] The official code repository for the paper "Mirage or Method? How Model–Task Alignment Induces Divergent RL Conclusions".☆19Feb 9, 2026Updated 6 months ago
- RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback☆30Mar 30, 2026Updated 5 months ago
- ☆23May 3, 2025Updated last year
- ☆15Feb 27, 2024Updated 2 years ago
- Code for EMNLP2023 paper "MolCA: Molecular Graph-Language Modeling with Cross-Modal Projector and Uni-Modal Adapter".☆13Dec 27, 2023Updated 2 years ago
- This is a repository for awesome any2any work collection.☆30Jul 10, 2026Updated last month
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [EMNLP 2022] TaCube: Pre-computing Data Cubes for Answering Numerical-Reasoning Questions over Tabular Data☆17May 17, 2023Updated 3 years ago
- ☆34Sep 19, 2025Updated 11 months ago
- A list of things to do in Montréal.☆29Oct 6, 2025Updated 10 months ago
- [ICLR 2026] RuleReasoner: Reinforced Rule-based Reasoning via Domain-aware Dynamic Sampling☆42Feb 25, 2026Updated 6 months ago
- TopoClaw: open-source cross-device AI agent for Android and Windows GUI automation, mobile-use, computer-use, social collaboration, proac…☆23May 12, 2026Updated 3 months ago
- DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents☆24Aug 4, 2025Updated last year
- [arXiv] "Linear Dynamics in the RLVR Training of Large Language Models"☆19May 25, 2026Updated 3 months ago
- Pushing Test-Time Scaling Limits of Deep Search with Asymmetric Verification☆22Oct 8, 2025Updated 10 months ago
- Evaluation utilities based on SymPy.☆25Dec 12, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Code-A1: Adversarial Evolving of Code LLM and Test LLM via Reinforcement Learning☆35Mar 17, 2026Updated 5 months ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 4 months ago
- [AAAI 2026] Test-Time Reinforcement Learning for GUI Grounding via Region Consistency https://arxiv.org/abs/2508.05615☆69Nov 8, 2025Updated 9 months ago
- [ACL 2025] Research code for the paper "OS-Kairos: Adaptive Interaction for MLLM-Powered GUI Agents"☆23Jun 19, 2025Updated last year
- [EMNLP 2025 Findings] Familiarity-aware Evidence Compression for Retrieval Augmented Generation☆15Aug 20, 2025Updated last year
- Official implementation of paper "Learning to Optimize Multi-objective Alignment Through Dynamic Reward Weighting"☆28Dec 31, 2025Updated 8 months ago
- ☆17Mar 9, 2026Updated 5 months ago