Code of Paper: Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models
☆16Mar 17, 2026Updated 4 months ago
Alternatives and similar repositories for ITP
Users that are interested in ITP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Aligning Agentic World Models via Knowledgeable Experience Learning☆37May 15, 2026Updated 2 months ago
- [NeurIPS D&B Track 2024] Source code for the paper "Constrained Human-AI Cooperation: An Inclusive Embodied Social Intelligence Challenge…☆25May 2, 2025Updated last year
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 4 months ago
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆66Apr 13, 2026Updated 3 months ago
- Codes for Mitigating Unhelpfulness in Emotional Support Conversations with Multifaceted AI Feedback (ACL 2024 Findings)☆17Jul 2, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for the paper "Self-Detoxifying Language Models via Toxification Reversal" (EMNLP 2023)☆18Oct 17, 2023Updated 2 years ago
- Improving transparency of large language models' reasoning☆15Nov 25, 2025Updated 7 months ago
- [ACL 2026] Enabling Efficient Reasoning in LLMs via Black-box Persuasive Prompting☆22Jan 9, 2026Updated 6 months ago
- Zero-WAM, an in-context world model for zero-shot robotic task generalization☆31Jul 8, 2026Updated last week
- [NeurIPS 2023] Bilevel Coreset Selection in Continual Learning: A New Formulation and Algorithm☆15Nov 23, 2023Updated 2 years ago
- "Parallel Test-Time Scaling for Latent Reasoning Models"☆22Apr 12, 2026Updated 3 months ago
- ☆29Aug 25, 2024Updated last year
- ICML 2024 - Self-Driven Entropy Aggregation for Byzantine-Robust Heterogeneous Federated Learning☆10Jul 16, 2024Updated 2 years ago
- ☆12Feb 28, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆40May 29, 2025Updated last year
- ☆12Jul 30, 2025Updated 11 months ago
- Github repository for "Internalizing World Models via Self-Play Finetuning for Agentic RL"☆36Nov 1, 2025Updated 8 months ago
- Official Implementation for the paper "Integrative Decoding: Improving Factuality via Implicit Self-consistency"☆33Apr 12, 2025Updated last year
- PyTorch implementation of experiments in the paper Aligning Language Models with Human Preferences via a Bayesian Approach☆32Nov 6, 2023Updated 2 years ago
- Official PyTorch implementation for the ICML 2023 paper "Out-of-Distribution Generalization of Federated Learning via Implicit Invariant …☆14Oct 31, 2023Updated 2 years ago
- ☆42Jun 11, 2025Updated last year
- Code Repository for NeurIPS 2021 accepted paper, named "Torwards Gradient-based Bilevel Optimization with non-convex Followers and Beyond…☆11Mar 28, 2022Updated 4 years ago
- ☆15Feb 25, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- !!!!(DEMO)!!!! !!! CHECK OUT THE NEW VERSİON !!! Counting Close People with Yolov7☆13Sep 14, 2022Updated 3 years ago
- [EMNLP 2025] WebAgent-R1: Training Web Agents via End-to-End Multi-Turn Reinforcement Learning☆94Nov 4, 2025Updated 8 months ago
- ☆36May 24, 2025Updated last year
- ☆13Apr 30, 2025Updated last year
- ICA Lens: compact ICA-based interpretability tools for exploring LLM activations. Code release for the paper.☆36Jul 5, 2026Updated 2 weeks ago
- 🤖 Code for our EMNLP 2022 paper: "BotsTalk: Machine-sourced Framework for Automatic Curation of Large-scale Multi-skill Dialogue Dataset…☆16Oct 7, 2024Updated last year
- A collection of long-term memory papers☆34Jan 18, 2026Updated 6 months ago
- Awesome Long-CoT Data☆22Mar 26, 2025Updated last year
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆15Jan 12, 2026Updated 6 months ago
- Official code for paper "SPA-RL: Reinforcing LLM Agent via Stepwise Progress Attribution"☆89Sep 13, 2025Updated 10 months ago
- ENACT is a benchmark that evaluates embodied cognition through world modeling from egocentric interaction. It is designed to be simple an…☆52Nov 27, 2025Updated 7 months ago
- PyTorch Implementation of Variance Reduced Optimization Algorithms -- SARAH and SVRG.☆15Jul 11, 2021Updated 5 years ago
- 使用Decoder-only的Transformer进行时序预测,包含SwiGLU和RoPE(Rotary Positional Embedding),Time series prediction using Decoder-only Transformer, Includ…☆16Jan 25, 2024Updated 2 years ago
- ☆17Oct 12, 2023Updated 2 years ago
- ☆19Feb 20, 2024Updated 2 years ago