Code for Arxiv 2023: Improving Language Model Negociation with Self-Play and In-Context Learning from AI Feedback
☆210May 24, 2023Updated 3 years ago
Alternatives and similar repositories for GPT-Bargaining
Users that are interested in GPT-Bargaining are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-agent Social Simulation + Efficient, Effective, and Stable alternative of RLHF. Code for the paper "Training Socially Aligned Langu…☆356Jun 18, 2023Updated 3 years ago
- ☆12Jul 4, 2024Updated 2 years ago
- ☆41Nov 30, 2023Updated 2 years ago
- DialOp: Decision-oriented dialogue environments for collaborative language agents☆114Nov 15, 2024Updated last year
- Repo for paper "CODIS: Benchmarking Context-Dependent Visual Comprehension for Multimodal Large Language Models".☆13Oct 14, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ChatArena (or Chat Arena) is a Multi-Agent Language Game Environments for LLMs. The goal is to develop communication and collaboration ca…☆1,563Aug 11, 2025Updated last year
- Benchmarking large language models' complex reasoning ability with chain-of-thought prompting☆2,775Aug 4, 2024Updated 2 years ago
- TART: A plug-and-play Transformer module for task-agnostic reasoning☆202Jun 22, 2023Updated 3 years ago
- The dataset and code for paper: TheoremQA: A Theorem-driven Question Answering dataset☆161Apr 23, 2024Updated 2 years ago
- The official repository for the paper "From Zero to Hero: Examining the Power of Symbolic Tasks in Instruction Tuning".☆64Apr 18, 2023Updated 3 years ago
- Repo for paper "Unleashing Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration"☆355May 8, 2024Updated 2 years ago
- Embroid: Unsupervised Prediction Smoothing Can Improve Few-Shot Classification☆11Aug 12, 2023Updated 3 years ago
- ICML 2024: Improving Factuality and Reasoning in Language Models through Multiagent Debate☆552Apr 24, 2025Updated last year
- ZYN: Zero-Shot Reward Models with Yes-No Questions☆34Aug 15, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Pytorch implementation for "Compressed Context Memory For Online Language Model Interaction" (ICLR'24)☆63Apr 18, 2024Updated 2 years ago
- Code for RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs. ACL 2023.☆63Nov 27, 2024Updated last year
- Resource, Evaluation and Detection Papers for ChatGPT☆456Mar 21, 2024Updated 2 years ago
- [NeurIPS 2023] PyTorch code for Can Language Models Teach? Teacher Explanations Improve Student Performance via Theory of Mind☆66Dec 21, 2023Updated 2 years ago
- Code accompanying the paper "R-U-SURE? Uncertainty-Aware Code Suggestions By Maximizing Utility Across Random User Intents"☆24Jul 8, 2026Updated 2 months ago
- Code for "Small Models are Valuable Plug-ins for Large Language Models"☆131May 16, 2023Updated 3 years ago
- Exchange-of-Thought: Enhancing Large Language Model Capabilities through Cross-Model Communication☆21Mar 21, 2024Updated 2 years ago
- Improving Language Understanding from Screenshots. Paper: https://arxiv.org/abs/2402.14073☆32Jul 9, 2024Updated 2 years ago
- Large-scale, Informative, and Diverse Multi-round Chat Data (and Models)☆2,903Mar 13, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆157Mar 18, 2023Updated 3 years ago
- Code for ACL2023 paper: Pre-Training to Learn in Context☆106Jul 26, 2024Updated 2 years ago
- This repo is to demo the concept of lossless compression with Transformers as encoder and decoder.☆14May 2, 2024Updated 2 years ago
- Codes and Data for Scaling Relationship on Learning Mathematical Reasoning with Large Language Models☆269Sep 12, 2024Updated 2 years ago
- [EMNLP 2023] Enabling Large Language Models to Generate Text with Citations. Paper: https://arxiv.org/abs/2305.14627☆527Oct 9, 2024Updated last year
- open-source code for paper: Retrieval Head Mechanistically Explains Long-Context Factuality☆241Aug 2, 2024Updated 2 years ago
- The Official Repository for "Bring Your Own Data! Self-Supervised Evaluation for Large Language Models"☆108Sep 23, 2023Updated 3 years ago
- Self-Alignment with Principle-Following Reward Models☆170Sep 18, 2025Updated last year
- Official repository for ACL 2025 paper "Model Extrapolation Expedites Alignment"☆75May 20, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- CodeUltraFeedback: aligning large language models to coding preferences (TOSEM 2025)☆76Jun 25, 2024Updated 2 years ago
- This is the official implementation of "Progressive-Hint Prompting Improves Reasoning in Large Language Models"☆208Oct 11, 2023Updated 2 years ago
- A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)☆3,751Feb 8, 2026Updated 7 months ago
- LLMs can generate feedback on their work, use it to improve the output, and repeat this process iteratively.☆822Oct 4, 2024Updated last year
- ☆135Jul 8, 2024Updated 2 years ago
- FireAct: Toward Language Agent Fine-tuning☆296Oct 22, 2023Updated 2 years ago
- This is the repo for the paper Shepherd -- A Critic for Language Model Generation☆225Aug 10, 2023Updated 3 years ago