A beamer template for LAMDA lab at NJU
☆16Oct 17, 2020Updated 5 years ago
Alternatives and similar repositories for LAMDA-Beamer-Template
Users that are interested in LAMDA-Beamer-Template are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Sep 14, 2020Updated 5 years ago
- [NeurIPS'20] Code for the paper "Offline Imitation Learning with a Misspecified Simulator"☆12Nov 24, 2021Updated 4 years ago
- ☆30Mar 1, 2022Updated 4 years ago
- RLA is a tool for managing your RL experiments automatically☆71Feb 7, 2023Updated 3 years ago
- The Official Code for Offline Model-based Adaptable Policy Learning (NeurIPS'21 & TPAMI)☆25Jan 16, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆10Oct 15, 2020Updated 5 years ago
- Toolkit of Causal Model-based Reinforcement Learning.☆33Jun 5, 2023Updated 3 years ago
- A python module designed for agile RL algorithm developing.☆26Jul 11, 2024Updated 2 years ago
- Re-implementations of SOTA RL algorithms.☆137Sep 7, 2023Updated 2 years ago
- A systematic design process for a self-organizing neuro-fuzzy Q-network for model-free and offline reinforcement learning.☆11May 29, 2023Updated 3 years ago
- RLA is a tool for managing your RL experiments automatically☆31Jan 11, 2025Updated last year
- Code for Paper (Policy Optimization in RLHF: The Impact of Out-of-preference Data)☆29Dec 19, 2023Updated 2 years ago
- ZOSVRG-BlackBox-Adv☆12Oct 30, 2018Updated 7 years ago
- Unofficial Code for NeurIPS 2021 paper "Regret Minimization Experience Replay in Off-policy Reinforcement Learning"☆14May 24, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆17Jan 24, 2024Updated 2 years ago
- ☆30Dec 22, 2022Updated 3 years ago
- Replicating Imagination-Augmented Agents for Deep Reinforcement Learning☆20Dec 17, 2017Updated 8 years ago
- GNOME Shell Extensions - Backup Tools / 备份工具☆16Jul 19, 2020Updated 6 years ago
- PyTorch implementation of Distribution Correction(DisCor) based on Soft Actor-Critic.☆37Jun 22, 2022Updated 4 years ago
- nju compilers☆19May 28, 2018Updated 8 years ago
- Kuaishou Online RL Benchmark☆19Oct 21, 2023Updated 2 years ago
- Benchmarked implementations of Offline RL Algorithms.☆77Mar 4, 2025Updated last year
- send and receive message and file by python3 socket☆12May 24, 2018Updated 8 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- 智源杯天文数据算法挑战赛 2020.1.15 - 2020.4.2☆12May 18, 2020Updated 6 years ago
- 🛠Robust SSH: auto-reconnect SSH session that preserves your running shell and command. Intuitive, no server-side setup, aimed at simplic…☆13Nov 14, 2025Updated 8 months ago
- Plannable Approximations to MDP Homomorphisms: Equivariance under Actions☆30Jun 30, 2020Updated 6 years ago
- ☆20Oct 27, 2025Updated 9 months ago
- ☆14Mar 24, 2023Updated 3 years ago
- d2l-zh(动手学深度学习)tensorflow2.0的代码实现.☆21Apr 25, 2022Updated 4 years ago
- A benchmark for evaluating reinforcement learning algorithms that train the policies using imaginary rollouts from LLMs.☆15Nov 4, 2025Updated 8 months ago
- [ICLR 22] Value Gradient weighted Model-Based Reinforcement Learning.☆25Apr 15, 2023Updated 3 years ago
- Imitation learning from multiple experts☆13Aug 29, 2022Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Code for Adapting Environment Sudden Changes by Learning Context Sensitive Policy☆21Jun 1, 2022Updated 4 years ago
- Code for ICLR 2022 Paper (HyperDQN: A Randomized Exploration Method for Deep Reinforcement Learning)☆12Nov 28, 2023Updated 2 years ago
- The code repository for "OmniEvalKit: A Modular, Lightweight Toolbox for Evaluating Large Language Model and its Omni-Extensions"☆13Feb 21, 2025Updated last year
- D3PE (Deep Data-Driven Policy Evaluation) aims to evaluation a large set of candidate policies from a fixed dataset to select best ones.☆10Jun 2, 2022Updated 4 years ago
- Algorithms described in the paper Hindsight Credit Assignment (NeurIPS 2019).☆11Oct 27, 2019Updated 6 years ago
- A repo containing bash scripts to deploy reinforcement learning dev environment within one click!☆11Jun 28, 2026Updated last month
- ☆12May 14, 2024Updated 2 years ago