Code for Mathematical Foundations of Reinforcement Learning
☆17Mar 31, 2025Updated last year
Alternatives and similar repositories for Code-for-Mathematical-Foundations-of-Reinforcement-Learning
Users that are interested in Code-for-Mathematical-Foundations-of-Reinforcement-Learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code for the LoG2022 paper -- MSGNN: A Spectral Graph Neural Network Based on a Novel Magnetic Signed Laplacian.☆14Updated this week
- Diffusion-based generative drug-like molecular editing with chemical natural language☆18Dec 22, 2024Updated last year
- Official implementation of "Learning To Draft: Adaptive Speculative Decoding with Reinforcement Learning" (ICLR 2026)☆23Mar 1, 2026Updated 6 months ago
- Unofficial implementation of Chain of Hindsight (https://arxiv.org/abs/2302.02676) using pytorch and huggingface Trainers.☆11Apr 5, 2023Updated 3 years ago
- Source code for AAAI2019 paper "Cash-out User Detection based on Attributed Heterogeneous Information Network with a Hierarchical Attenti…☆15Nov 12, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆12Jun 19, 2018Updated 8 years ago
- underwater dataset, open-data☆12Aug 22, 2021Updated 5 years ago
- Official code for "Automated Scoring for Reading Comprehension via In-context BERT Tuning" (AIED 2022)☆13May 23, 2022Updated 4 years ago
- A list of academic papers that utilize the ManiSkill framework, categorized by year of publication.☆25May 30, 2025Updated last year
- Official implementation of Stackelberg PPO for morphology–control co-design.☆18Mar 17, 2026Updated 5 months ago
- This is official code implementation of the <Revisiting Neural Networks for Continual Learning: An Architectural Perspective> in IJCAI 20…☆13Nov 25, 2024Updated last year
- Genetic-guided GFlowNets for Sample Efficient Molecular Optimization (NeurIPS 2024)☆21Nov 14, 2025Updated 9 months ago
- 2020 PTA History Test Questions☆15Feb 7, 2023Updated 3 years ago
- Graph generative pre-trained transformer☆22Jun 3, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Repository to store all necessary packages for common manipulator integrations☆21Nov 29, 2023Updated 2 years ago
- Double-Ended Synthesis Planning with Goal-Constrained Bidirectional Search (NeurIPS 2024)☆31Jan 23, 2025Updated last year
- A benchmark for evaluating future-event forecasting from audio and video context in multimodal language models☆29Jan 22, 2026Updated 7 months ago
- This repository includes code for our paper: ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning…☆15May 2, 2026Updated 4 months ago
- Harvard AM205: group activity files☆18Nov 1, 2021Updated 4 years ago
- Associated paper: https://arxiv.org/abs/2010.04971☆14Jan 2, 2022Updated 4 years ago
- 考研政治自制刷题GUI☆15Dec 7, 2023Updated 2 years ago
- Shanghai Droid Robot CO.,LTD☆13Dec 22, 2024Updated last year
- [ACL 2024] ReactXT: Understanding Molecular “Reaction-ship” via Reaction-Contextualized Molecule-Text Pretraining. by Zhiyuan Liu*, Yaoru…☆31Sep 3, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Sccheduling Environment for Multi-Robot Coordination Problems☆19May 9, 2022Updated 4 years ago
- ☆36Jan 20, 2026Updated 7 months ago
- ICLR 2026☆16Apr 29, 2026Updated 4 months ago
- ☆31Jul 21, 2026Updated last month
- Official implementation for "How Should We Meta-Learn Reinforcement Learning Algorithms?"☆23Sep 7, 2025Updated last year
- My own version from "Writing a C Compiler" Book from NoStarchPress using C++ and LLVM libraries.☆41May 18, 2026Updated 3 months ago
- Project page for the RSS 2025 paper Resolving Conflicting Constraints in Multi-Agent Reinforcement Learning with Layered Safety☆16Aug 7, 2025Updated last year
- Python wrapper for Lambda Lanczos☆17May 10, 2026Updated 3 months ago
- NeurIPS24: Aligning Target-Aware Molecule Diffusion Models with Exact Energy Optimization☆45Apr 2, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Multi-Task Multi-Agent Reinforcement Learning via Skill Graphs☆17Jul 19, 2025Updated last year
- A tiny script which makes methods of the `Math` object available to numbers by adding properties to `Number.prototype`☆16Jun 26, 2015Updated 11 years ago
- ☆25Oct 22, 2025Updated 10 months ago
- TensorFlow code and pre-trained models for BERT☆12Mar 19, 2019Updated 7 years ago
- Implementing the base algorithms of Deep Reinforcement Learning in Python☆22Sep 30, 2023Updated 2 years ago
- My notes life-full-stack☆21Apr 8, 2021Updated 5 years ago
- I2Q: A Fully Decentralized Q-Learning Algorithm☆19Nov 10, 2022Updated 3 years ago