☆19Oct 30, 2025Updated 10 months ago
Alternatives and similar repositories for KnapsackRL
Users that are interested in KnapsackRL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for ICLR 2022 Paper (HyperDQN: A Randomized Exploration Method for Deep Reinforcement Learning)☆12Nov 28, 2023Updated 2 years ago
- Code for Paper (Policy Optimization in RLHF: The Impact of Out-of-preference Data)☆29Dec 19, 2023Updated 2 years ago
- This repo contains the code for the paper "Understanding and Mitigating Hallucinations in Large Vision-Language Models via Modular Attrib…☆39Jul 14, 2025Updated last year
- Official code for "Algorithmic Capabilities of Random Transformers" (NeurIPS 2024)☆16Sep 28, 2024Updated last year
- Irene is a python package that aims to be a toolkit for global optimization problems that can be realized algebraically. It generalizes L…☆15Sep 18, 2026Updated last week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code for Paper (ReMax: A Simple, Efficient and Effective Reinforcement Learning Method for Aligning Large Language Models)☆201Dec 16, 2023Updated 2 years ago
- Source for the sample efficient tabular RL submission to the 2019 NIPS workshop on Biological and Artificial RL☆25Apr 14, 2022Updated 4 years ago
- ☆45Mar 22, 2024Updated 2 years ago
- ☆10Jul 13, 2024Updated 2 years ago
- Efficient parallelizable algorithms for multidimensional arrays to speed up your data pipelines☆22Jan 28, 2026Updated 7 months ago
- Code for Adam-mini: Use Fewer Learning Rates To Gain More https://arxiv.org/abs/2406.16793☆460May 13, 2025Updated last year
- Simple script for running interactive masked language model with pre-trained BERT models.☆18May 3, 2020Updated 6 years ago
- ☆18Sep 24, 2024Updated 2 years ago
- A library for datasets containing heterogeneous data☆13Oct 3, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Code for the paper: Why Transformers Need Adam: A Hessian Perspective☆65Mar 11, 2025Updated last year
- Implementation of Action Matching for the Schrödinger equation☆25Jun 18, 2023Updated 3 years ago
- ☆14May 30, 2019Updated 7 years ago
- This project applies Monte Carlo Tree Search (MCTS) to a simple grid world.☆10May 30, 2018Updated 8 years ago
- ☆20Dec 5, 2022Updated 3 years ago
- sc14 matlab application☆14Nov 24, 2014Updated 11 years ago
- TD-VAE in PyTorch☆10May 28, 2019Updated 7 years ago
- ☆13Apr 3, 2019Updated 7 years ago
- ☆15Feb 22, 2018Updated 8 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- AISTATS 2019: Reference-based Adversarial Sampling & Its applications to Soft Q-learning☆15Jan 21, 2019Updated 7 years ago
- ☆11Jan 12, 2021Updated 5 years ago
- MIPT course☆17Apr 15, 2021Updated 5 years ago
- Ant Gather and Ant Maze envs, separated from RLLab☆11Aug 2, 2018Updated 8 years ago
- 16824 homework: weakly supervised object detection with PyTorch☆13Sep 5, 2018Updated 8 years ago
- noiseprint2 is a porting of noiseprint to tensorflow 2 and keras☆12Feb 20, 2021Updated 5 years ago
- Reproduction of the complete process of DeepSeek-R1 on small-scale models, including Pre-training, SFT, and RL.☆33Mar 11, 2025Updated last year
- ☆15Jan 9, 2026Updated 8 months ago
- JOYTOU is a BootStrap blog template developed by Joytou Wu.☆10Feb 5, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Levin tree search guided by both a policy and a heuristic function☆19Jul 13, 2023Updated 3 years ago
- Question Answering In Context☆28Nov 24, 2022Updated 3 years ago
- Exploration Strategies for Deep Reinforcement Learning☆39Oct 31, 2018Updated 7 years ago
- python algorithms to solve sparse linear programming problems☆34Jul 6, 2023Updated 3 years ago
- 我们是第一个完全可商用的角色大模型。☆39Aug 11, 2024Updated 2 years ago
- Using PyTorch autograd to compute Hessian of Perplexity for Large Language Models☆29Apr 17, 2025Updated last year
- Reinforcement learning tutorials using the rlberry library.☆18Jan 9, 2023Updated 3 years ago