Reinforcement learning - Batched Impala - PyTorch - Mario Kart
☆13Jul 21, 2020Updated 6 years ago
Alternatives and similar repositories for Batched-Impala-PyTorch
Users that are interested in Batched-Impala-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Related papers for offline reforcement learning (we mainly focus on representation and sequence modeling and conventional offline RL)☆19Apr 21, 2022Updated 4 years ago
- Structured Object-Aware Physics Prediction for Video Modeling and Planning☆32May 9, 2020Updated 6 years ago
- Gym implementation of connector to Deepmind lab☆13Mar 26, 2019Updated 7 years ago
- Represented Value Function Approach for Large Scale Multi Agent Reinforcement Learning☆17Mar 11, 2020Updated 6 years ago
- Python implementation of Gibbs sampling for the naı̈ve Bayes model presented by Resnik and Hardisty☆14Feb 10, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Original tensorflow implementation of SILOT (Spatially Invariant, Label-free Object Tracking).☆13Mar 24, 2023Updated 3 years ago
- A cog model for the all-mpnet-base-v2 sentence-transformers embedding model.☆15Jan 3, 2024Updated 2 years ago
- Deployed version of Tableaunoir. Do not modify this repository.☆11Apr 19, 2026Updated 4 months ago
- (ICLR 2021) Learning to Represent Action Values as a Hypergraph on the Action Vertices☆23Jun 22, 2021Updated 5 years ago
- ☆16Jan 22, 2018Updated 8 years ago
- Pytorch Implementation of the Distributed SAC. Test environment is LunarLanderContinuous-v2 and Metaworld MT1, MT10☆12Apr 6, 2022Updated 4 years ago
- Library for Auto-Encoding Sequential Monte Carlo☆18Jan 29, 2024Updated 2 years ago
- Source code for student lectures on dependent type theory.☆12Jun 9, 2025Updated last year
- Neural Fictitious Self-Play in Leduc Holdem☆11Jul 4, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Reinforcement Learning Algorithms with Unity 3D Environments☆18Jul 15, 2019Updated 7 years ago
- Implementation of Direct Preference Optimization☆17Jul 17, 2023Updated 3 years ago
- Docker containers of baseline agents for the Crafter environment☆30Dec 14, 2021Updated 4 years ago
- Representation Learning in RL☆13Jun 1, 2022Updated 4 years ago
- Codes accompanying the paper "Context-Aware Sparse Deep Coordination Graphs (https://arxiv.org/abs/2106.02886).☆21Mar 15, 2022Updated 4 years ago
- A visualizer for Docker Swarm using the Docker Remote API, Node.JS, and D3☆12Sep 30, 2016Updated 9 years ago
- The source code of the Documentation of the Asymptote Geometry Module☆10May 15, 2026Updated 3 months ago
- Terminal Program to do SVG to ASCII art☆14Oct 3, 2020Updated 5 years ago
- Julia interface for Gradescope autograding☆10Dec 24, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Scraper for PhET Science & Math Interactive Simulations☆20Updated this week
- Implementation of Sum-Product Attend-Infer-Repeat☆31May 21, 2020Updated 6 years ago
- ML's radishal Universal Levenshtein Automata library.☆13Jan 5, 2022Updated 4 years ago
- [ICLR 2024] Closing the Gap between TD Learning and Supervised Learning - A Generalisation Point of View.☆25Apr 19, 2024Updated 2 years ago
- A versatile AI chatbot leveraging function-calling language models via Ollama. Features include advanced function calling, self-reflectio…☆17Feb 1, 2025Updated last year
- ☆11Oct 22, 2020Updated 5 years ago
- Script for processing OpenAI's PRM800K process supervision dataset into an Alpaca-style instruction-response format☆27Jul 12, 2023Updated 3 years ago
- Javascript implementation of Fractran☆15Sep 14, 2017Updated 8 years ago
- Setup of my personal infrastructure. sweet☆15Dec 2, 2025Updated 8 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code for phonetically classifying TIMIT using TensorFlow☆17Jul 1, 2016Updated 10 years ago
- [Python] [arXiv/cs] Paper "An Overview of Gradient Descent Optimization Algorithms" by Sebastian Ruder☆24Mar 23, 2019Updated 7 years ago
- ☆13Jan 30, 2017Updated 9 years ago
- SMASH: Physics-guided Reconstruction of Collisions from Videos, SIGGRAPH Asia 2016☆11Jan 25, 2018Updated 8 years ago
- ☆12Aug 30, 2021Updated 5 years ago
- Display Clock within your REPL☆18Mar 27, 2022Updated 4 years ago
- Jupyter notebook extension that creates a canvas of a cell and allows to paint onto the contents so it could be used to explain some conc…☆13Jul 17, 2019Updated 7 years ago