🐲 Stanford CS234 : Reinforcement Learning
☆27Jun 8, 2019Updated 7 years ago
Alternatives and similar repositories for CS234_RL
Users that are interested in CS234_RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Stanford CS234 : Reinforcement Learning☆194Oct 3, 2019Updated 6 years ago
- Stanford CS234: Reinforcement Learning Winter 2020☆19Mar 24, 2023Updated 3 years ago
- Minimal example to access PyBullet using C++☆12Mar 19, 2021Updated 5 years ago
- Reinforcement Learning Seminar at the Chinese University of Hong Kong, Shenzhen, China.☆21Nov 17, 2023Updated 2 years ago
- a benchmark to evaluate the situated inductive reasoning☆20Jan 7, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- My Solutions of Assignments of CS234: Reinforcement Learning Winter 2019☆170Mar 24, 2023Updated 3 years ago
- Prompt Optimization with Human Feedback☆18Aug 7, 2024Updated 2 years ago
- Code repository accompanying the CHI 2021 Paper titled "Adapting User Interfaces with Model-based Reinforcement Learning"☆17Oct 18, 2021Updated 4 years ago
- This repository contains my solution to the Stanford Course cs224u "Natural Language Understanding" Summer 2019☆12Nov 7, 2019Updated 6 years ago
- A prerender demo for Vue 3 base on Vite.☆10Jun 2, 2022Updated 4 years ago
- Teleop Twist Keyboard for ROS2☆26Oct 29, 2023Updated 2 years ago
- Implementation of "Learning Across Tasks and Domains" ICCV 2019☆15Mar 24, 2023Updated 3 years ago
- OptiDICE: Offline Policy Optimization via Stationary Distribution Correction Estimation☆16Aug 3, 2023Updated 3 years ago
- Framework for autonomous vehicle risk assessment☆16May 11, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Benchmark result of different RL algorithms on MetaDrive environments, including Multi-agent RL (IPPO, centralized critics, CoPO).☆16Oct 25, 2022Updated 3 years ago
- ☆11Jan 12, 2015Updated 11 years ago
- Solutions to coding assignments of Stanford Reinforcement Learning course Winter 2021☆13Aug 29, 2021Updated 5 years ago
- Compiled lecture notes for AA274☆23Mar 22, 2019Updated 7 years ago
- ☆27Apr 22, 2024Updated 2 years ago
- Gitlet 自己实现的本地Git版本控制工具(Berkeley CS61B 数据结构课程项目)☆15Oct 7, 2023Updated 2 years ago
- Reward shaping approach for instruction following settings, leveraging language at multiple levels of abstraction.☆21Mar 9, 2021Updated 5 years ago
- Github Pages template for academic personal websites, forked from mmistakes/minimal-mistakes☆11Feb 22, 2024Updated 2 years ago
- A workspace for learning computer science and software engineering topics☆13Oct 31, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- C++ implementation of the algorithm in "Fast and Accurate Least-Mean-Squares Solvers", NIPS19☆11Mar 4, 2020Updated 6 years ago
- ☆11Aug 26, 2023Updated 3 years ago
- Code to reproduce the experiments in The Mirage of Action-Dependent Baselines in Reinforcement Learning.☆17Aug 2, 2018Updated 8 years ago
- Scalable sim-to-real transfer of soft robot designs☆16Jun 27, 2020Updated 6 years ago
- Urho3D extra minimal examples and demos. Tested in Ubuntu 18.04.☆11Feb 25, 2022Updated 4 years ago
- ☆12Sep 15, 2026Updated last week
- Variants for ROS (implemented as metapackages)☆11May 31, 2025Updated last year
- Natural Language to SQL☆13Oct 13, 2023Updated 2 years ago
- Tock Tracker (a tock is like a pomodoro but longer)☆11Jul 28, 2014Updated 12 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- SLAHAN is an implementation of Kamigaito et al., 2020, "Syntactically Look-A-Head Attention Network for Sentence Compression", In Proc. o…☆17Jan 27, 2021Updated 5 years ago
- Auxiliary variable Markov chain Monte Carlo methods☆10Oct 24, 2017Updated 8 years ago
- ☆10Jun 29, 2021Updated 5 years ago
- ☆11Aug 1, 2019Updated 7 years ago
- Landing a Spaceship using Upside-Down Reinforcement Learning (a.k.a ⅂ꓤ)☆13Oct 25, 2023Updated 2 years ago
- Sequence models in Numpy☆25Oct 9, 2020Updated 5 years ago
- Linear Algebra for Machine Learning Book Exercises☆13May 19, 2019Updated 7 years ago