Reproduce ICLR2018 submission "Emergent Communication through Negotiation"
☆18Apr 19, 2018Updated 8 years ago
Alternatives and similar repositories for emergent-comms-negotiation
Users that are interested in emergent-comms-negotiation are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reproducing the reinforcement learning models used in "Emergence of Linguistic Communication from Referential Games with Symbolic and Pix…☆12Jun 23, 2018Updated 8 years ago
- An implementation of Emergence of Grounded Compositional Language in Multi-Agent Populations by Igor Mordatch and Pieter Abbeel☆78Feb 21, 2018Updated 8 years ago
- ☆16Oct 23, 2023Updated 2 years ago
- Emergent Communication Pretraining for Few-Shot Machine Translation☆13Dec 3, 2020Updated 5 years ago
- Code for reproducing the results from the paper Avoiding Side Effects in Complex Environments☆12Jun 3, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code repository for On the interaction between supervision and self-play in emergent communication (ICLR 2020)☆15Feb 4, 2020Updated 6 years ago
- Implementation of the Hierarchical and Interpretable Skill Acquisition in Multi-task Reinforcement Learning by Tianmin Shu, Caiming Xiong…☆11Jun 18, 2018Updated 8 years ago
- Source code for "Influencing Long-Term Behavior in Multiagent Reinforcement Learning" (NeurIPS 2022)☆19Jan 1, 2023Updated 3 years ago
- ☆10Feb 22, 2018Updated 8 years ago
- ROS and LCM drivers for OptiTrack's Motive 2 software. Optimized for tracking aerial drones. Runs on Ubuntu Linux.☆22Jul 28, 2020Updated 6 years ago
- Code release for Learning with Opponent-Learning Awareness and variations.☆156Apr 13, 2023Updated 3 years ago
- ☆20Oct 31, 2025Updated 11 months ago
- Repository for code experimenting with RL and Solar Tracking☆13Apr 24, 2018Updated 8 years ago
- This is a sample implementation of "TIMERS: Error-Bounded SVD Restart on Dynamic Networks"(AAAI 2018).☆12Jul 4, 2018Updated 8 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Co-Adaptation of Algorithmic and Implementational Innovations in Inference-based Deep Reinforcement Learning (NeurIPS2021)☆20Oct 25, 2021Updated 4 years ago
- ☆10Sep 14, 2023Updated 3 years ago
- Code for the paper "TD or not TD: Analyzing the Role of Temporal Differencing in Deep Reinforcement Learning", Artemij Amiranashvili, Ale…☆12Aug 24, 2018Updated 8 years ago
- Source code for "A Policy Gradient Algorithm for Learning to Learn in Multiagent Reinforcement Learning" (ICML 2021)☆34Oct 6, 2022Updated 4 years ago
- Pytorch implementation of Stable Opponent Shaping (https://openreview.net/pdf?id=SyGjjsC5tQ).☆22Jan 15, 2020Updated 6 years ago
- Repo for reproduction of sequential social dilemmas☆416Mar 6, 2025Updated last year
- Train and Visualize Binary Neural Networks (Code for: The High-Dimensional Geometry of Binary Neural Networks)☆13Jan 31, 2018Updated 8 years ago
- Code release to our paper on an agent-based model of the Ramsey-Cass-Koopmans macroeconomic model. In this model, the households imitate …☆13Jun 3, 2021Updated 5 years ago
- Matlab code for learning doubly sparse dictionary on synthetic data. Details can be found in the paper "A Provable Approach for Double-Sp…☆11Mar 5, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- On the pitfalls of measuring emergent communication☆35Mar 12, 2019Updated 7 years ago
- ☆22Dec 8, 2022Updated 3 years ago
- Contains an implementation of "Imitation Learning via Kernel Mean Embedding (2018, AAAI)"☆11Oct 2, 2018Updated 8 years ago
- Round 1 Starter Kit for the MarLo challenge☆21Sep 27, 2018Updated 8 years ago
- growing interpretable part graphs on convnets via multi-shot learning, in AAAI 2017☆15May 28, 2017Updated 9 years ago
- PickTime Chrome Extension - extract myvisit tokens and send to PickTime bot☆13May 16, 2022Updated 4 years ago
- ☆23Jan 25, 2023Updated 3 years ago
- Obsidian Plugin to execute squiggle in a note.☆26Sep 25, 2022Updated 4 years ago
- [EMNLP 2017] Code for "Natural Language Does Not Emerge 'Naturally' in Multi-Agent Dialog"☆95May 5, 2020Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Github Repo for CARL: Cautious Adaptation for RL in Safety Critical Settings☆14Nov 22, 2022Updated 3 years ago
- we propose a novel and efficient cross-domain human parsing model to bridge the cross-domain differences in terms of visual appearance an…☆15Jan 9, 2018Updated 8 years ago
- A Reinforcement learning model which applies deterministic policy gradient algorithms to maximize return on portfolio management task☆14Jun 2, 2018Updated 8 years ago
- ☆16Mar 2, 2019Updated 7 years ago
- Tensorflow code for WACV 2019 paper "Attention Based Natural Language Grounding by Navigating Virtual Environment" - https://arxiv.org/ab…☆17Nov 7, 2018Updated 7 years ago
- Collection of awesome Call for Papers to submit your Reinforcement Learning papers☆18Sep 13, 2021Updated 5 years ago
- Parallel implementation of the ridge detection algorithm for curve reconstruction in CUDA☆13Nov 21, 2017Updated 8 years ago