A Collection for Distributed Reinforcement Learning Papers
☆18Sep 24, 2025Updated 11 months ago
Alternatives and similar repositories for Awesome-Distributed-RL
Users that are interested in Awesome-Distributed-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PiKV: KV Cache Management System for Mixture of Experts [Efficient ML System]☆63Aug 17, 2026Updated 2 weeks ago
- FastCache: Fast Caching for Diffusion Transformer Through Learnable Linear Approximation [Efficient ML Model]☆52Aug 2, 2026Updated 3 weeks ago
- A Serving System for Distributed and Parallel LLM Quantization [Efficient ML System]☆26Jun 18, 2025Updated last year
- DuoDecoding: Hardware-aware Heterogeneous Speculative Decoding with Dynamic Multi-Sequence Drafting☆19Mar 4, 2025Updated last year
- Accelerating Multitask Training Trough Adaptive Transition [Efficient ML Model]☆12May 23, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- tmp DPI☆14Dec 18, 2024Updated last year
- AVPipe :-)☆12Jul 16, 2021Updated 5 years ago
- Serpens is an HBM FPGA accelerator for SpMV☆23Jul 26, 2024Updated 2 years ago
- ☆13Aug 14, 2026Updated 2 weeks ago
- ☆11Jul 28, 2021Updated 5 years ago
- [CVPR 2025] PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models☆54Jun 12, 2025Updated last year
- This is pytorch implmentation project of Bootsrapped DQN☆13Dec 6, 2020Updated 5 years ago
- Adaptive Topology Reconstruction for Robust Graph Representation Learning [Efficient ML Model]☆10Feb 11, 2025Updated last year
- Generative Quantum Circuits☆64Apr 28, 2026Updated 4 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [ACL 2026 (Main)] LongSpec: Long-Context Lossless Speculative Decoding with Efficient Drafting and Verification☆86Jul 14, 2025Updated last year
- Pytorch implementation of our paper accepted by ICML 2024 -- CaM: Cache Merging for Memory-efficient LLMs Inference☆51Jun 19, 2024Updated 2 years ago
- ☆49Mar 3, 2024Updated 2 years ago
- LLM-Aided FPGA Design for Signal Processing Applications☆35Jun 4, 2025Updated last year
- The official implementation of paper: SimLayerKV: A Simple Framework for Layer-Level KV Cache Reduction.☆55Oct 18, 2024Updated last year
- Study on the application of reinforcement learning to the management of a traffic light intersection.☆14Mar 6, 2023Updated 3 years ago
- codes for paper 《Neighborhood Cooperative Multiagent Reinforcement Learning for Adaptive Traffic Signal Control in Epidemic Regions》☆14Apr 3, 2022Updated 4 years ago
- Code for paper "Successor Uncertainties: Exploration and Uncertainty in Temporal Difference Learning" by David Janz*, Jiri Hron*, Przemys…☆22Feb 24, 2023Updated 3 years ago
- ☆10May 20, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Prompt format and padding guide for Llama 2☆12Sep 18, 2023Updated 2 years ago
- C++17 implementation of einops for libtorch - clear and reliable tensor manipulations with einstein-like notation