Minimal code for A Generalist Agent
☆44Nov 4, 2022Updated 3 years ago
Alternatives and similar repositories for Gato-A-Generalist-Agent
Users that are interested in Gato-A-Generalist-Agent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Unofficial Gato: A Generalist Agent☆220Jan 14, 2024Updated 2 years ago
- Implementation of GATO style Generalist Multimodal model capable of image, text, RL and Robotics tasks☆45Jun 19, 2024Updated 2 years ago
- ☆18Jul 10, 2022Updated 4 years ago
- Codebase for ICLR 2023 paper, "SMART: Self-supervised Multi-task pretrAining with contRol Transformers"☆53Jan 26, 2024Updated 2 years ago
- panda_gym integration to use an AI to move the real robot☆11Apr 14, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The AI Arena: A framework for distributed multi-agent reinforcement learning☆14Aug 5, 2022Updated 4 years ago
- Single-file truly minimal implementation of state-of-the-art reinforcement learning algorithms.☆21Feb 13, 2023Updated 3 years ago
- Semi-Supervised Offline Reinforcement Learning with Action-Free Trajectories☆43Jul 16, 2023Updated 3 years ago
- Implementation of ICML 2023 paper: Future-conditioned Unsupervised Pretraining for Decision Transformer☆29Jul 25, 2023Updated 3 years ago
- Experiments to train transformer network to master reinforcement learning environments.☆32Mar 14, 2021Updated 5 years ago
- Code for NeurIPS paper "Self-Organized Group for Cooperative Multi-agentReinforcement Learning".☆22Feb 20, 2023Updated 3 years ago
- The official implementation of "Transformer in Transformer as Backbone for Deep Reinforcement Learning"☆58Dec 27, 2023Updated 2 years ago
- A videogame made with PyGame turned into an Open AI Gym Learning Environment for Reinforcement Learning agents.☆14Jan 3, 2023Updated 3 years ago
- Implementation of "Do As I Can, Not As I Say: Grounding Language in Robotic Affordances" by Google☆26Aug 3, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is the code of reproducing the results of our paper: On the importance of Hyperparameter Optimization for Model-based Reinforcement …☆16Aug 19, 2021Updated 4 years ago
- ☆12Jan 30, 2021Updated 5 years ago
- ☆12Mar 15, 2022Updated 4 years ago
- Semi-Markov Afterstate Actor-Critic (SMAAC) with Maze☆11Nov 16, 2021Updated 4 years ago
- Decision Transformer for offline single-agent autonomous highway driving☆28Jun 19, 2023Updated 3 years ago
- RLFP (CoRL 2024)☆14Oct 11, 2024Updated last year
- ☆11Jan 11, 2022Updated 4 years ago
- Improving upon state of the art cooperative deep reinforcement learning in StarCraft II☆13May 16, 2019Updated 7 years ago
- Boosts JupyterLab - proving tools to help you organise, explore and analyse your experimental data.☆16Apr 24, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆13May 7, 2023Updated 3 years ago
- ☆17Oct 12, 2023Updated 2 years ago
- A disaggregated memory orchestration system that virtualizes cluster wide memory to scale data intensive, large memory workloads in virtu…☆13Apr 26, 2019Updated 7 years ago
- Combining Evolutionary Algorithms and deep Reinforcement Learning☆19Jul 17, 2018Updated 8 years ago
- Posted at AAAI 2023☆11Sep 4, 2025Updated 11 months ago
- a suite of finetuned LLMs for atomically precise function calling 🧪☆16Aug 10, 2026Updated last week
- Code used in our paper "Robust Deep Reinforment Learning through Adversarial Loss"☆33Oct 3, 2023Updated 2 years ago
- a moderate and simulated space non-cooperative object visual tracking dataset, which contains 60 binocular video sequences with manual an…☆12Sep 25, 2021Updated 4 years ago
- Tiktok is an advanced multimedia recommender system that fuses the generative modality-aware collaborative self-augmentation and contrast…☆14Aug 18, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆19Mar 1, 2023Updated 3 years ago
- ☆10Jan 19, 2022Updated 4 years ago
- ☆11Mar 18, 2021Updated 5 years ago
- Official code for ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning (AAAI'24)☆17Feb 10, 2024Updated 2 years ago
- finetuning shakespeare on karpathy/nanoGPT☆23Feb 2, 2023Updated 3 years ago
- A bridge between ROS and Blender☆25Oct 6, 2014Updated 11 years ago
- Multi-task Multi-agent Soft Actor Critic for SMAC☆15Jan 18, 2022Updated 4 years ago