Github repository for "Internalizing World Models via Self-Play Finetuning for Agentic RL"
☆37Nov 1, 2025Updated 10 months ago
Alternatives and similar repositories for SPA
Users that are interested in SPA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2026] Aligning Agentic World Models via Knowledgeable Experience Learning☆40May 15, 2026Updated 3 months ago
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models☆22Apr 14, 2026Updated 4 months ago
- [ICML 2025] Closed-Loop Long-Horizon Robotic Planning via Equilibrium Sequence Modeling☆13May 5, 2025Updated last year
- AEGIS: Automated Error Generation and Attribution for Multi-Agent Systems☆26Aug 17, 2026Updated 2 weeks ago
- [ACL'24 Oral] Analysing The Impact of Sequence Composition on Language Model Pre-Training☆24Aug 18, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆19Jun 2, 2026Updated 3 months ago
- We introduce Reasoning via Video, a new paradigm that uses maze-solving video generation to probe multimodal reasoning; our VR-Bench show…☆65Feb 4, 2026Updated 6 months ago
- [ACL 2026 SAC Highlight Award] Can We Predict Before Executing Machine Learning Agents?☆24Jul 7, 2026Updated last month
- [EMNLP main 2026]Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models☆16Mar 17, 2026Updated 5 months ago
- [ICLR 2023, ICLR DG oral] PAIR, the optimizer and model selection criteria for OOD Generalization☆54Apr 12, 2024Updated 2 years ago
- Code for Evolving Language Models without Labels: Majority Drives Selection, Novelty Promotes Variation (EVOL-RL).☆51Mar 31, 2026Updated 5 months ago
- ENACT is a benchmark that evaluates embodied cognition through world modeling from egocentric interaction. It is designed to be simple an…☆53Nov 27, 2025Updated 9 months ago
- A multi-agent framework to help with your homework.☆11Mar 1, 2025Updated last year
- Scaling Agentic Environments Automatically.☆68Mar 26, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆16Mar 3, 2024Updated 2 years ago
- ☆33Jan 7, 2025Updated last year
- ☆12Jun 5, 2024Updated 2 years ago
- Official code for paper "SPA-RL: Reinforcing LLM Agent via Stepwise Progress Attribution"☆93Sep 13, 2025Updated 11 months ago
- Official Project Page for Web World Models (https://arxiv.org/abs/2512.23676)☆92Jun 5, 2026Updated 2 months ago
- Evaluate Multimodal LLMs as Embodied Agents☆60Feb 14, 2025Updated last year
- Code for paper 'Are We Falling in a Middle-Intelligence Trap? An Analysis and Mitigation of the Reversal Curse'☆14Aug 2, 2024Updated 2 years ago
- World model reinforcement learning for multi-turn VLM agents. RL for vision framework (NeurIPS 2025).☆494Updated this week
- ReCAP: Recursive Context-Aware Reasoning and Planning for Large Language Model Agents, NeurIPS 2025☆42Nov 15, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Reasoning-based Evaluation and Ranking of Translations.☆22Jun 2, 2026Updated 2 months ago
- A machine learning based fine-grained disease transmission simulator☆12Apr 20, 2020Updated 6 years ago
- [NeurIPS D&B Track 2024] Source code for the paper "Constrained Human-AI Cooperation: An Inclusive Embodied Social Intelligence Challenge…☆25May 2, 2025Updated last year
- R1V, trained with AI feedback, answers open-ended visual questions.☆14Apr 12, 2025Updated last year
- ☆11Jun 28, 2019Updated 7 years ago
- ☆14Oct 22, 2024Updated last year
- Automatic Recall Machines: Internal Replay, Continual Learning and the Brain☆11Jul 14, 2020Updated 6 years ago
- In‑Context World‑Action Modeling from Human Videos for Open‑Ended Task Generalization☆41Updated this week
- ☆116Jan 8, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Mitigating Spurious Correlations in Multi-modal Models during Fine-tuning (ICML 2023)☆19Dec 15, 2023Updated 2 years ago
- PyTorch codes for the paper "An Empirical Study of Multimodal Model Merging"☆37Oct 11, 2023Updated 2 years ago
- ☆11Oct 25, 2024Updated last year
- Official data and code for the paper "VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents".☆14Mar 18, 2026Updated 5 months ago
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 5 months ago
- ☆26Apr 11, 2023Updated 3 years ago
- [CVPR 2023] "TrojViT: Trojan Insertion in Vision Transformers" by Mengxin Zheng, Qian Lou, Lei Jiang☆15Jan 5, 2024Updated 2 years ago