A distributed GPU-centric experience replay system for large AI models.
☆19Aug 1, 2023Updated 3 years ago
Alternatives and similar repositories for gear
Users that are interested in gear are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Mar 5, 2024Updated 2 years ago
- Exploring techniques to generate diverse conventions in multi-agent settings☆16Nov 14, 2023Updated 2 years ago
- A python package to design and debug RL agents.☆35Apr 2, 2026Updated 5 months ago
- Overcooked human-AI experiment platform☆41Dec 21, 2023Updated 2 years ago
- ☆144Jan 30, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Tempo is a system for declarative, efficient, end-to-end compiled dynamic deep learning☆31Oct 21, 2025Updated 10 months ago
- Supporting code for "Learning to Solve Combinatorial Graph Partitioning Problems via Efficient Exploration".☆13Jun 18, 2022Updated 4 years ago
- IJCAI 2019 - Regularized Opponent Model with Maximum Entropy Objective (ROMMEO)☆23Dec 8, 2022Updated 3 years ago
- ☆15Apr 11, 2024Updated 2 years ago
- ☆14Nov 28, 2023Updated 2 years ago
- we're building an AI to play the board game Diplomacy!☆36Mar 27, 2022Updated 4 years ago
- ☆30Aug 20, 2021Updated 5 years ago
- A Multi-agent Learning Framework☆62May 10, 2021Updated 5 years ago
- A markdown-supported command-line interface tool that connects to ChatGPT using OpenAI's API key.☆48May 29, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆17Apr 14, 2024Updated 2 years ago
- Repository for running LLMs efficiently on Mac silicon (M1, M2, M3). Features Jupyter notebook for Meta-Llama-3 setup using MLX framework…☆11May 4, 2024Updated 2 years ago
- Overcooked-AI Experiment Psiturk Demo (for MTurk experiments)☆13May 10, 2021Updated 5 years ago
- ☆19May 4, 2023Updated 3 years ago
- Official Repo for "SplitQuant / LLM-PQ: Resource-Efficient LLM Offline Serving on Heterogeneous GPUs via Phase-Aware Model Partition and …☆39Aug 29, 2025Updated last year
- Adaptation of DQN, DDQN and COMA for multi-agent Gym environments☆10Oct 3, 2023Updated 2 years ago
- A large-scale multi-modal pre-trained model☆134Feb 7, 2023Updated 3 years ago
- 来记录一波 pybind11 实例~☆18Nov 19, 2022Updated 3 years ago
- The repository for 'Unsupervised Learning for Combinatorial Optimization with Principled Proxy Design'☆16Oct 9, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Pathfinding Using Reinforcement Learning☆12May 21, 2019Updated 7 years ago
- ☆45Jan 9, 2024Updated 2 years ago
- ☆10Apr 23, 2021Updated 5 years ago
- ☆16Feb 20, 2024Updated 2 years ago
- 大二上学期--计算机组成与设计(PH)--实验☆15Apr 23, 2019Updated 7 years ago
- ☆17Dec 4, 2019Updated 6 years ago
- This is the code of CoCo-MILP: Inter-Variable Contrastive and Intra-Constraint Competitive MILP Solution Prediction. AAAI 2026 Oral.☆16May 13, 2026Updated 3 months ago
- Official resporitory for "IPDPS' 24 QSync: Quantization-Minimized Synchronous Distributed Training Across Hybrid Devices".☆20Feb 23, 2024Updated 2 years ago
- AI model training on heterogeneous, geo-distributed resources☆46Nov 24, 2025Updated 9 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Solve the advection diffusion equations looped into an optimization problem with JAX/autodiff☆14May 8, 2025Updated last year
- A Simple, Distributed and Asynchronous Multi-Agent Reinforcement Learning Framework for Google Research Football AI.☆120Jan 16, 2024Updated 2 years ago
- ☆12Jan 30, 2021Updated 5 years ago
- Recurrent Network-based Deterministic Policy Gradient for Solving Bipedal Walking Challenge on Rugged Terrains☆12Oct 16, 2017Updated 8 years ago
- A lightweight RL environment for query optimization.☆15Sep 13, 2024Updated last year
- develop 1) an on-path DNS packet injector, and 2) a passive DNS poisoning attack detector. Part 1: The DNS packet injector you are goin…☆18Apr 21, 2018Updated 8 years ago
- The first, open access evaluation dataset for methods to identify bias by word choice and labeling☆26Oct 30, 2025Updated 10 months ago