Actor Critic model to play Cartpole game
☆52Aug 4, 2018Updated 8 years ago
Alternatives and similar repositories for Actor-Critic-pytorch
Users that are interested in Actor-Critic-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 原稿用紙;原稿紙;稿紙;日式便箋;UPTEX/UPLATEX 縱書☆10Nov 27, 2019Updated 6 years ago
- Implementation of DeepAR in PyTorch.☆10Aug 5, 2019Updated 7 years ago
- Accompanying code for our NeurIPS 2019 paper☆11Nov 7, 2019Updated 6 years ago
- Learning Transferable Features with Deep Adaptation Networks☆12Jul 18, 2023Updated 3 years ago
- Simple model for sentence compression (a.k.a Baseline in Klerke et al., NAACL 2016)☆10Dec 16, 2018Updated 7 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A toy example of Policy Gradient implemented in Pytorch☆95Jan 24, 2018Updated 8 years ago
- This is the paddle code for SeBoW(Self-Born wiring for neural trees), a kind of neural tree born form a large search space☆11Dec 10, 2021Updated 4 years ago
- An implementation of effective policy ensemble.☆16Jul 5, 2023Updated 3 years ago
- ☆24Feb 22, 2023Updated 3 years ago
- ☆13Sep 8, 2024Updated 2 years ago
- [TMLR 2025] A collection of research papers on constraint inference within the field of RL☆11May 9, 2025Updated last year
- ☆11Jul 20, 2023Updated 3 years ago
- Code for Colangelo and Lee (2025)☆17Feb 3, 2025Updated last year
- pix2pix and Cycle GAN architectures for image style transfer☆13May 27, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆14Mar 5, 2024Updated 2 years ago
- ☆10Dec 21, 2024Updated last year
- 《自然语言处理——基于预训练模型的方法》全书代码实现☆12Jan 16, 2023Updated 3 years ago
- ☆13Sep 19, 2023Updated 3 years ago
- An attempt to reverse engineer custom file formats used by the game Outlaws from LucasArts.☆16Aug 23, 2026Updated last month
- PyTorch implementation of Advantage Actor-Critic (A2C)☆47Nov 25, 2017Updated 8 years ago
- ☆19Apr 15, 2024Updated 2 years ago
- The MATLAB source code☆15Dec 2, 2019Updated 6 years ago
- code for sentence compression☆20Mar 3, 2018Updated 8 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 🎾 Multi-Agent Proximal Policy Optimization approach to a competitive reinforcement learning problem☆22Sep 25, 2022Updated 4 years ago
- One implementation of the paper "Controllable Neural Dialogue Summarization with Personal Named Entity Planning" (EMNLP 2022).☆18Nov 9, 2023Updated 2 years ago
- This code is for cross-domain segmentation tasks, which can plot T-sne of domains and classes☆18May 9, 2023Updated 3 years ago
- Solving the OpenAI Gym (MountainCarContinuous-v0) with DDPG☆21Jan 23, 2023Updated 3 years ago
- ☆13May 25, 2017Updated 9 years ago
- ☆10Mar 24, 2023Updated 3 years ago
- Lernd is ∂ILP (dILP) framework implementation based on Deepmind's paper Learning Explanatory Rules from Noisy Data.☆27Mar 25, 2023Updated 3 years ago
- Experiments codes for RecSys '21 paper "Mitigating Confounding Bias in Recommendation via Information Bottleneck"☆19Apr 6, 2022Updated 4 years ago
- Train neural networks to use as SMC and importance sampling proposals☆24Dec 6, 2017Updated 8 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 一个南开大学beamer模板☆15May 21, 2021Updated 5 years ago
- Re-produce DQN, REINFORCE, REINFORCE with baseline, one-step AC, QAC, QAC with shared network, PPO2, DDPG, TD3, SAC, SAC discrete,A2C,A3C☆21Jul 27, 2020Updated 6 years ago
- ☁️ KUMO: Generative Evaluation of Complex Reasoning in Large Language Models☆21Jun 4, 2025Updated last year
- ☆18Jul 25, 2024Updated 2 years ago
- Code for CoRL 2019 paper "TuneNet: One-Shot Residual Tuning for System Identification and Sim-to-Real Robot Task Transfer"☆24Mar 31, 2020Updated 6 years ago
- ☆10Jul 5, 2023Updated 3 years ago
- Network simulator for edge computing and cloud computing☆24Sep 26, 2017Updated 9 years ago