🤖 Implements of Reinforcement Learning algorithms.
☆117Apr 1, 2018Updated 8 years ago
Alternatives and similar repositories for Reinforcement-Learning
Users that are interested in Reinforcement-Learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 💸 Papers and Code Implements for Quantitative-Trading☆41May 9, 2018Updated 8 years ago
- 📈 Personae is a repo of implements and environment of Deep Reinforcement Learning & Supervised Learning for Quantitative Trading.☆1,410Nov 29, 2018Updated 7 years ago
- 将 DQN 应用在微信跳一跳小程序☆14Feb 4, 2018Updated 8 years ago
- Fixed-point arithmetic in C++☆12Sep 25, 2013Updated 12 years ago
- Generating Families of Practical Fast Matrix Multiplication Algorithms☆12Jul 7, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code for the ICRA2018 paper "Learning with training wheels: Speeding up training with a simple controller for Deep Reinforcement Learning…☆38Dec 4, 2018Updated 7 years ago
- This project contains several Deep Reinforcement Learning method and some experiments basd on OpenAi gym.☆19Jan 28, 2018Updated 8 years ago
- Implementation of Structural Correspondence Learning☆15Apr 22, 2018Updated 8 years ago
- Implementation of selected reinforcement learning algorithms in Tensorflow. A3C, DDPG, REINFORCE, DQN, etc.☆153May 28, 2023Updated 3 years ago
- [ICRA 2021] Learning Robot Trajectories subject to Kinematic Joint Constraints☆12Jul 29, 2025Updated last year
- Multiple object tracking using Kalman filters and Munkres algorithm☆13Jun 7, 2017Updated 9 years ago
- Augmentation scripts for the bAbI Dialog Tasks dataset☆13Oct 16, 2018Updated 7 years ago
- Python library used to communicate and control Kuka manipulator (tested on KR16 model). Under continuous development.☆12Aug 9, 2018Updated 8 years ago
- Waste Sorting with Robot Arm Tossing☆10Sep 19, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A path planning framework based on Sampling-based algorithm and Deep Reinforcement learning.☆10May 9, 2023Updated 3 years ago
- This is a tuned sparse matrix dense vector multiplication(SpMV) library☆23Mar 21, 2016Updated 10 years ago
- ☆16Feb 22, 2024Updated 2 years ago
- Autonomous Driving on Carla simulator using Deep Deterministic Policy Gradients. Based on Kendall, et. al. 2018.☆13Apr 2, 2019Updated 7 years ago
- A Jensen-Shannon Divergence Driven Mechanistic Study of Context Attribution in Retrieval-Augmented Generation☆16Aug 28, 2025Updated last year
- soft q learning and soft actor critic☆16Dec 23, 2018Updated 7 years ago
- A Structured Self-attentive Sentence Embedding Zhouhan Lin, Minwei Feng, Cicero Nogueira dos Santos, Mo Yu, Bing Xiang, Bowen Zhou, Yoshu…☆11Nov 23, 2017Updated 8 years ago
- Strassen's Algorithm for Tensor Contraction☆15Jul 7, 2017Updated 9 years ago
- 机器学习和量化分析学习进行中☆380Feb 3, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PyTorch bindings for openai-gemm☆20Feb 6, 2017Updated 9 years ago
- Catamount is a compute graph analysis tool to load, construct, and modify deep learning models and to symbolically analyze their compute …☆14May 18, 2021Updated 5 years ago
- Deep reinforcement learning baselines base on OpenAI. More algorithms are included, such as Rainbow: Combining Improvements in Deep Rei…☆35Aug 23, 2018Updated 8 years ago
- Setting up DDPG based reinforcement learning in ROS Gazebo environment☆14Jul 29, 2019Updated 7 years ago
- ☆13Apr 7, 2025Updated last year
- LibCP -- A Library for Conformal Prediction☆13Feb 26, 2015Updated 11 years ago
- DRL-based collision avoidance for turtlebot3☆19Feb 6, 2023Updated 3 years ago
- "JABAS: Joint Adaptive Batching and Automatic Scaling for DNN Training on Heterogeneous GPUs" (EuroSys '25)☆16Apr 7, 2025Updated last year
- Clinical NLP concept extraction of ADEs in the 2018 n2c2 Adverse Drug Events and Medication Extraction (Track 2). Includes data preproce…☆16Nov 21, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Graph AI generates neurological hypotheses validated in molecular, organoid, and clinical systems☆35Dec 27, 2025Updated 8 months ago
- Parallel implementation of k-means clustering using MPI4PY and PyCUDA.☆10Mar 11, 2019Updated 7 years ago
- TensorFlow implementation of FAIR's InferSent (Supervised Learning of Universal Sentence Representations from Natural Language Inference …☆14Aug 6, 2018Updated 8 years ago
- A repository for the Analytics Working Group.☆22Dec 20, 2022Updated 3 years ago
- It contains Data Augmentaion, Strided convolution, Batch Normalization, Leaky Relu, Global Average pooling, L2 Regularization, learning …☆12Jun 3, 2018Updated 8 years ago
- Code base, outputs, scoring guidelines and annotations for ChatGPT NCCN evaluation | Paper: https://jamanetwork.com/journals/jamaoncology…☆18Aug 27, 2023Updated 3 years ago
- Reinforcement learning algorithm implementations and ML experimentation workspace☆45Jun 8, 2019Updated 7 years ago