This is the official implementation of the paper "DipLLM: Fine-Tuning LLM for Strategic Decision-making in Diplomacy".
☆26Dec 19, 2025Updated 9 months ago
Alternatives and similar repositories for dipllm
Users that are interested in dipllm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repo supports integrating LLMs and communication algorithms with MARL using SMAC as the platform. It provides an end-to-end workflow…☆21Mar 8, 2025Updated last year
- ☆40Jan 3, 2025Updated last year
- ☆15Jan 24, 2025Updated last year
- suPER is a collaborative multi-agent RL algorithm☆14Jun 11, 2024Updated 2 years ago
- ☆59Oct 21, 2025Updated 10 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- MineRL DDPG Agent to Obtain Diamond in Minecraft☆14Jan 21, 2020Updated 6 years ago
- This is the official PyTorch implementation of the paper "Boosting Continuous Control with Consistency Policy".☆48Nov 11, 2025Updated 10 months ago
- Tufts Probabilistic Robotics Spring 2020 Final Project☆17May 7, 2020Updated 6 years ago
- (ICRA 2025) Inverse Mixed Strategy Games with Generative Trajectory Models☆18May 21, 2025Updated last year
- Code for the paper "Reactive Exploration to Cope with Non-Stationarity in Lifelong Reinforcement Learning"☆16Jul 4, 2022Updated 4 years ago
- code repo for EMNLP'21 Finding Counter-Interference Adapter for Multilingual Machine Translation☆18Oct 19, 2022Updated 3 years ago
- This Python script performs a Model Predictive Control (MPC) simulation for vehicle lateral control using the CasADi framework. The main …☆19Feb 18, 2025Updated last year
- RA-L 2023: Learning to Play Trajectory Games Against Opponents with Unknown Objectives: A differentiable adaptive game-theoretic planner …☆28Sep 10, 2024Updated 2 years ago
- 6-DOF Spacecraft Orbit and Attitude Simulation (Comes broken since I had to rewrite quickly & remove sensitive stuff before exporting saf…☆18Jan 17, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆18Jul 22, 2025Updated last year
- Skill Design From AI Feedback☆33Feb 27, 2025Updated last year
- Benchmark for LLMs playing full press diplomacy☆65Mar 4, 2025Updated last year
- ☆16Oct 21, 2025Updated 10 months ago
- Improving Math reasoning through Direct Preference Optimization with Verifiable Pairs☆21Mar 20, 2025Updated last year
- ☆12Feb 27, 2022Updated 4 years ago
- Code accompanying paper "Models as Agents: Optimizing Multi-Step Predictions of Interactive Local Models in Model-Based Multi-Agent Reinf…☆15Dec 2, 2023Updated 2 years ago
- ☆12Oct 7, 2020Updated 5 years ago
- Advantage Alignment Algorithms (ICLR 2025 oral)☆19Apr 7, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 2021 Reddit r/roguelikedev summer tutorial series☆12Jun 7, 2026Updated 3 months ago
- ☆25Oct 11, 2019Updated 6 years ago
- WasmNoise, Fast Noise Generation For The Web☆12Jan 1, 2020Updated 6 years ago
- ☆21Mar 19, 2024Updated 2 years ago
- Official code release for ICLR23 "Diminishing Return of Value Expansion Methods in Model-Based Reinforcement Learning"☆16Mar 8, 2023Updated 3 years ago
- JsonWebToken implementation for cairo-lang http://self-issued.info/docs/draft-ietf-oauth-json-web-token.html☆13May 30, 2022Updated 4 years ago
- Code for the paper LaM-SLidE - Latent Space Modeling of Spatial Dynamical Systems via Linked Entities☆26May 23, 2025Updated last year
- A gate level simulator and gate level netlist standard specification in Cairo☆17Feb 15, 2022Updated 4 years ago
- [ECCV 2024] 💐Official implementation of the paper "Diffusion Reward: Learning Rewards via Conditional Video Diffusion"☆121Jul 2, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR'26] MARSHAL: Incentivizing Multi-Agent Reasoning via Self-Play with Strategic LLMs☆57Apr 17, 2026Updated 5 months ago
- Official code for "Decoding-Time Language Model Alignment with Multiple Objectives".☆30Oct 30, 2024Updated last year
- Find the samples, in the test data, on which your (generative) model makes mistakes.☆32Oct 16, 2024Updated last year
- ☆22Oct 4, 2021Updated 4 years ago
- ☆23Oct 4, 2021Updated 4 years ago
- Multi-Agent RL for UAV Control for Fair and Energy-Efficient Coverage Maximisation☆21May 18, 2025Updated last year
- Understanding how features learned by neural networks evolve throughout training☆40Oct 24, 2024Updated last year