Paper notes for my PhD on Machine Learning (mostly focused on Reinforcement Learning)
☆17Jul 22, 2019Updated 7 years ago
Alternatives and similar repositories for paper_notes
Users that are interested in paper_notes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository contains implementations of the paper VUSFA☆14Mar 31, 2021Updated 5 years ago
- Decoupling Dynamics and Reward for Transfer Learning☆16Sep 7, 2018Updated 7 years ago
- Labs for understanding and coding Standard Reinforcement Learning concepts☆60Jan 17, 2019Updated 7 years ago
- ☆17Jun 30, 2022Updated 4 years ago
- Online demo of DRLViz, an interactive tool to understand decisions and memory in Deep Reinforcement Learning☆16Dec 8, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for AAAI 2023 paper "Hypernetworks for Zero-shot Transfer in Reinforcement Learning"☆24Apr 26, 2023Updated 3 years ago
- Multi-Agent training using Deep Deterministic Policy Gradient Networks, Solving the Tennis Environment☆11Oct 20, 2018Updated 7 years ago
- A package for fast evaluation of multivariate polynomials.☆13Oct 15, 2021Updated 4 years ago
- ☆20Aug 16, 2021Updated 4 years ago
- Imagination Augmented Agents TensorFlow☆26Mar 30, 2020Updated 6 years ago
- Code for abstracting, evaluating, and visualizing Markov Decision Processes.☆10Jan 12, 2017Updated 9 years ago
- ☆12Mar 23, 2018Updated 8 years ago
- Round 1 Starter Kit for the MarLo challenge☆21Sep 27, 2018Updated 7 years ago
- E2C implementation in PyTorch☆43Jul 5, 2017Updated 9 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Pytorch implementation of SCAN: Learning Abstract Hierarchical Compositional Visual Concepts☆20Jan 27, 2018Updated 8 years ago
- Learning bisimulation metrics for control, particularly suited to sparse reward settings☆11Feb 28, 2023Updated 3 years ago
- ☆13Apr 11, 2022Updated 4 years ago
- Reproduction of the paper "Soft Q-Learning with Mutual Information Regularization" CoRL 2019.☆10Jan 10, 2019Updated 7 years ago
- Incorporating Neuro-Inspired Adaptability for Continual Learning in Artificial Intelligence☆29Dec 12, 2023Updated 2 years ago
- Local search for NAS☆18Nov 3, 2020Updated 5 years ago
- Sources for algorithm selection for combinatorial search problems survey☆17Jul 10, 2019Updated 7 years ago
- Fast evaluation of multivariate polynomials☆18Jun 26, 2023Updated 3 years ago
- Code for Transformers are Adaptable Task Planners, CoRL 2022☆12Mar 28, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is the codebase for our ICRA 2020 submission, GraphRQI: Classifying Driver Behaviors Using Graph Spectrums.☆13Dec 8, 2019Updated 6 years ago
- POMDP formulation of a pedestrian avoidance problem for autonomous driving☆51Apr 3, 2020Updated 6 years ago
- Package implements a number local outlier factor algorithms for outlier detection and finding anomalous data☆12Jun 7, 2017Updated 9 years ago
- ☆10Oct 11, 2022Updated 3 years ago
- ☆10Jul 5, 2016Updated 10 years ago
- Python library implementing recommender systems algorithms with http://tensorflow.org☆12Dec 21, 2018Updated 7 years ago
- ☆11Sep 11, 2023Updated 2 years ago
- Useful scripts for iOS Pythonista app.☆10Apr 8, 2024Updated 2 years ago
- 股票高频数据(数据来源:新浪)☆13Jan 29, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- My documented journey to learning fastapi☆10Apr 30, 2023Updated 3 years ago
- ☆21May 13, 2019Updated 7 years ago
- Count based exploration with the successor representation for Unity ML's Pyramid☆12Jun 19, 2019Updated 7 years ago
- ☆10Apr 7, 2021Updated 5 years ago
- A generative model of compositionality in symmetric monoidal (Kleisli) categories☆12Oct 4, 2023Updated 2 years ago
- DiDi-Udacity Self-Driving Car Challenge 2017 Raw Data Reader☆11Apr 17, 2017Updated 9 years ago
- papers about reinforcement learning☆13Jan 4, 2021Updated 5 years ago