This is the official implementation of the paper "DipLLM: Fine-Tuning LLM for Strategic Decision-making in Diplomacy".
☆25Dec 19, 2025Updated 7 months ago
Alternatives and similar repositories for dipllm
Users that are interested in dipllm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repo supports integrating LLMs and communication algorithms with MARL using SMAC as the platform. It provides an end-to-end workflow…☆20Mar 8, 2025Updated last year
- Supervised and RL Models for No Press Diplomacy☆76Dec 4, 2025Updated 8 months ago
- ☆40Jan 3, 2025Updated last year
- ☆15Feb 26, 2026Updated 5 months ago
- ☆15Jan 24, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- suPER is a collaborative multi-agent RL algorithm☆14Jun 11, 2024Updated 2 years ago
- ☆56Oct 21, 2025Updated 9 months ago
- Tufts Probabilistic Robotics Spring 2020 Final Project☆17May 7, 2020Updated 6 years ago
- (ICRA 2025) Inverse Mixed Strategy Games with Generative Trajectory Models☆18May 21, 2025Updated last year
- code repo for EMNLP'21 Finding Counter-Interference Adapter for Multilingual Machine Translation☆18Oct 19, 2022Updated 3 years ago
- ☆11Nov 16, 2025Updated 8 months ago
- RA-L 2023: Learning to Play Trajectory Games Against Opponents with Unknown Objectives: A differentiable adaptive game-theoretic planner …☆27Sep 10, 2024Updated last year
- An OpenAI Gym environment for Pokemon battles☆12Sep 3, 2019Updated 6 years ago
- Skill Design From AI Feedback☆33Feb 27, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A Reinforcement Learning to play StarCraft, written in PyTorch. Currently in research.☆11Mar 23, 2018Updated 8 years ago
- ☆12Oct 7, 2020Updated 5 years ago
- Official Implementation of "RTop-K: Ultra-Fast Row-Wise Top-K Selection for Neural Network Acceleration on GPUs"☆28Jul 23, 2025Updated last year
- ☆25Oct 11, 2019Updated 6 years ago
- ☆25Jan 28, 2025Updated last year
- Official code release for ICLR23 "Diminishing Return of Value Expansion Methods in Model-Based Reinforcement Learning"☆16Mar 8, 2023Updated 3 years ago
- 中南大学智能科学与技术专业机器学习课程设计,其中包含自己实现的神经网络框架,可实现的模型有:ResNet,VGG16☆10Jul 6, 2022Updated 4 years ago
- [T-AI 2025] "TrafficGamer: Reliable and Flexible Traffic Simulation for Safety-Critical Scenarios with Game-Theoretic Oracles" official r…☆35Mar 16, 2026Updated 4 months ago
- 中南大学本科生毕业设计论文模板☆23Apr 1, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ECCV 2024] 💐Official implementation of the paper "Diffusion Reward: Learning Rewards via Conditional Video Diffusion"☆121Jul 2, 2024Updated 2 years ago
- FlockGPT is a novel approach to drone flocking control using natural language and generative AI. It features an LLM-based interface for u…☆22Aug 12, 2024Updated last year
- Quantification of Uncertainty with Adversarial Models☆29Jul 11, 2023Updated 3 years ago
- [ICLR'26] MARSHAL: Incentivizing Multi-Agent Reasoning via Self-Play with Strategic LLMs☆55Apr 17, 2026Updated 3 months ago
- Official code for "Decoding-Time Language Model Alignment with Multiple Objectives".☆30Oct 30, 2024Updated last year
- Find the samples, in the test data, on which your (generative) model makes mistakes.☆32Oct 16, 2024Updated last year
- Understanding how features learned by neural networks evolve throughout training☆41Oct 24, 2024Updated last year
- Released codes of the RETIA model.☆19Mar 20, 2024Updated 2 years ago
- Learning to Incentivize Other Learning Agents☆36Jun 13, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICLR 2023] Choreographer: a world-model-based agent that discovers and learns unsupervised skills in latent imagination, and it's able t…☆42Jun 18, 2024Updated 2 years ago
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.☆106Jul 19, 2025Updated last year
- Official codebase for "The Generalization Gap in Offline Reinforcement Learning" accepted to ICLR 2024☆29Apr 8, 2026Updated 4 months ago
- SDLG is an efficient method to accurately estimate aleatoric semantic uncertainty in LLMs☆28Jun 7, 2024Updated 2 years ago
- Retrieval-Augmented Decision Transformer: External Memory for In-context RL☆26Oct 27, 2024Updated last year
- ☆129May 30, 2023Updated 3 years ago
- [ICLR 2021] Group Equivariant Generative Adversarial Networks.☆14May 6, 2021Updated 5 years ago