Learning From Human Preferences - Tensorflow+Keras Implementation
☆18Aug 17, 2017Updated 9 years ago
Alternatives and similar repositories for LearningFromHumanPreferences
Users that are interested in LearningFromHumanPreferences are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reproduction of OpenAI and DeepMind's "Deep Reinforcement Learning from Human Preferences"☆338Nov 29, 2021Updated 4 years ago
- Code for Deep RL from Human Preferences [Christiano et al]. Plus a webapp for collecting human feedback☆565Jan 24, 2023Updated 3 years ago
- A simple moving dot environment for OpenAI Gym to test reinforcement learning algorithms☆23Sep 1, 2022Updated 4 years ago
- Code to reproduce Supervised Policy Update (ICLR 2019)☆17Dec 8, 2022Updated 3 years ago
- ☆19Mar 28, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A wrapper around MoveIt that enables more traditional industrial robot programming.☆15Jan 29, 2020Updated 6 years ago
- Inferring beliefs about dynamics from behavior☆30May 24, 2018Updated 8 years ago
- Deep Learning the Sorting Algorithm☆12Dec 11, 2016Updated 9 years ago
- C++ PyTorch Examples☆10Aug 18, 2019Updated 7 years ago
- (Experimental) ROS packages for Blue + Gazebo☆15Aug 4, 2019Updated 7 years ago
- ☆11Oct 6, 2020Updated 5 years ago
- Implementation of the paper <Model-based Reinforcement Learning for Predictions and Control for Limit Order Books (Wei et al., J.P. Morga…☆12Aug 22, 2023Updated 3 years ago
- ☆13Oct 31, 2021Updated 4 years ago
- ☆23Feb 18, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The Group Marching Tree Algorithm, presented at IRC 2017☆24Jun 13, 2017Updated 9 years ago
- A reinforcement learning package implemented in Torch☆11Jan 24, 2016Updated 10 years ago
- Notes for my Calculus courses in college, written in Jupyter Notebooks☆12Jul 31, 2016Updated 10 years ago
- Reproduction of OpenAI and DeepMind's "Deep Reinforcement Learning from Human Preferences"☆31Jul 27, 2021Updated 5 years ago
- My Data Provider: A minimal multi-exchange data providing project to feed trading algorithms/bots. Built with Python and FastAPI.☆11May 30, 2024Updated 2 years ago
- A trading system in python with GUI extension in PYQT. Proposed accepted API : many including those in README.☆11Jun 10, 2020Updated 6 years ago
- Open AI gym environment for the Baxter robot☆14Oct 6, 2016Updated 9 years ago
- A powerful text cleaner for Japanese web texts☆12Jan 20, 2024Updated 2 years ago
- open-source Mandarian biased word dataset☆14Sep 21, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Quadrotor control using deep reinforcement learning☆17Dec 14, 2017Updated 8 years ago
- Material from M1P1, formalised in Lean☆15Nov 2, 2019Updated 6 years ago
- ☯️ AllenNLP training configurations for promising models on Named Entity Recognition. (BiLSTM-CRF, BiLSTM-CNN-CRF, BERT, BERT-CRF)☆15Nov 26, 2020Updated 5 years ago
- https://mth229.github.io☆13Apr 20, 2026Updated 5 months ago
- Tutorials for the Robotics MVA 2023 class☆11Aug 1, 2024Updated 2 years ago
- ☆13Mar 2, 2018Updated 8 years ago
- A news based stock scalper using LLM and quant approach☆15Jan 16, 2025Updated last year
- A set of Deep Reinforcement Learning Agents implemented in Tensorflow.☆13Feb 5, 2017Updated 9 years ago
- Serverless Scraper for Cryptocurrency Order Book Data☆15Dec 8, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Probabilistic Motion Primives library☆14Dec 14, 2022Updated 3 years ago
- Prototyping mujoco simulation environments.☆11Feb 20, 2025Updated last year
- Library to compute kinematics, inverse kinematics and dynamics for N-dof robotics systems based on Denavit-Hartenberg convention.☆18Oct 28, 2025Updated 10 months ago
- ☆30Dec 2, 2021Updated 4 years ago
- Public accompanying repository for Universite de Montreal's IFT 6757: Autnonomous Vehicles, Fall 2019.☆11Jun 21, 2022Updated 4 years ago
- ☆14Mar 6, 2018Updated 8 years ago
- Accompanying code for the RSS 2019 paper, "Learning Reward Functions by Integrating Human Demonstrations and Preferences"☆12May 20, 2019Updated 7 years ago