[EMNLP main 2026]Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models
☆16Mar 17, 2026Updated 5 months ago
Alternatives and similar repositories for ITP
Users that are interested in ITP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2025] Closed-Loop Long-Horizon Robotic Planning via Equilibrium Sequence Modeling☆13May 5, 2025Updated last year
- [EMNLP 2026] Aligning Agentic World Models via Knowledgeable Experience Learning☆40May 15, 2026Updated 3 months ago
- [NeurIPS D&B Track 2024] Source code for the paper "Constrained Human-AI Cooperation: An Inclusive Embodied Social Intelligence Challenge…☆25May 2, 2025Updated last year
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 5 months ago
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models☆22Apr 14, 2026Updated 4 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆19Jun 2, 2026Updated 2 months ago
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆71Apr 13, 2026Updated 4 months ago
- Codes for Mitigating Unhelpfulness in Emotional Support Conversations with Multifaceted AI Feedback (ACL 2024 Findings)☆17Jul 2, 2024Updated 2 years ago
- [ACL 2026] Enabling Efficient Reasoning in LLMs via Black-box Persuasive Prompting☆22Jan 9, 2026Updated 7 months ago
- In‑Context World‑Action Modeling from Human Videos for Open‑Ended Task Generalization☆41Updated this week
- "Parallel Test-Time Scaling for Latent Reasoning Models"☆24Apr 12, 2026Updated 4 months ago
- ICML 2024 - Self-Driven Entropy Aggregation for Byzantine-Robust Heterogeneous Federated Learning☆10Jul 16, 2024Updated 2 years ago
- ☆12Feb 28, 2025Updated last year
- ☆40May 29, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆12Jul 30, 2025Updated last year
- Github repository for "Internalizing World Models via Self-Play Finetuning for Agentic RL"☆37Nov 1, 2025Updated 9 months ago
- Official Implementation for the paper "Integrative Decoding: Improving Factuality via Implicit Self-consistency"☆33Apr 12, 2025Updated last year
- PyTorch implementation of experiments in the paper Aligning Language Models with Human Preferences via a Bayesian Approach☆32Nov 6, 2023Updated 2 years ago
- Improving transparency of large language models' reasoning☆15Nov 25, 2025Updated 9 months ago
- OfHSV project using VGG16 and siamese neural network. Very easy. 使用VGG16和孪生神经网络实现离线签名鉴别,极其简单☆13Jan 23, 2024Updated 2 years ago
- ☆42Jun 11, 2025Updated last year
- !!!!(DEMO)!!!! !!! CHECK OUT THE NEW VERSİON !!! Counting Close People with Yolov7☆13Sep 14, 2022Updated 3 years ago
- ☆15Feb 25, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [EMNLP 2025] WebAgent-R1: Training Web Agents via End-to-End Multi-Turn Reinforcement Learning☆97Nov 4, 2025Updated 9 months ago
- ☆13May 13, 2021Updated 5 years ago
- ☆13Apr 30, 2025Updated last year
- ☆30Aug 25, 2024Updated 2 years ago
- 🤖 Code for our EMNLP 2022 paper: "BotsTalk: Machine-sourced Framework for Automatic Curation of Large-scale Multi-skill Dialogue Dataset…☆16Oct 7, 2024Updated last year
- See the official code and checkpoints for "Timer: Generative Pre-trained Transformers Are Large Time Series Models"☆16Aug 19, 2024Updated 2 years ago
- A collection of long-term memory papers☆37Jan 18, 2026Updated 7 months ago
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated 2 years ago
- Official code for paper "SPA-RL: Reinforcing LLM Agent via Stepwise Progress Attribution"☆93Sep 13, 2025Updated 11 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆15Jan 12, 2026Updated 7 months ago
- ENACT is a benchmark that evaluates embodied cognition through world modeling from egocentric interaction. It is designed to be simple an…☆53Nov 27, 2025Updated 9 months ago
- This repository contains the source code for the research presented in the paper "Exploring hidden flow structures from sparse data throu…☆12Dec 4, 2023Updated 2 years ago
- Math 228A 2019 Fall☆16Dec 4, 2019Updated 6 years ago
- Example Code for paper "Provably Faster Algorithms for Bilevel Optimization"☆15Dec 28, 2021Updated 4 years ago
- PyTorch Implementation of Variance Reduced Optimization Algorithms -- SARAH and SVRG.☆15Jul 11, 2021Updated 5 years ago
- [ACL 2025] RADAR: Enhancing Radiology Report Generation with Supplementary Knowledge Injection☆36Jul 23, 2025Updated last year