[QReward] RewardService Python Client, make RL Training reward function more faster
☆16Apr 23, 2026Updated 4 months ago
Alternatives and similar repositories for QReward
Users that are interested in QReward are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MrlX: A Multi-Agent Reinforcement Learning Framework☆221Jan 19, 2026Updated 8 months ago
- 国科大研究生课程 操作系统高级教程2023年思考题☆12Dec 24, 2023Updated 2 years ago
- My love.☆28Apr 1, 2026Updated 5 months ago
- An interface to program any congestion control protocol for an unreliable connection based protocol sent over UDP. It comes with a clean …☆12Apr 8, 2022Updated 4 years ago
- [TBD] "m4: A Learned Flow-level Network Simulator" by Chenning Li, Anton A. Zabreyko, Om Chabra, Arash Nasr-Esfahany, Kevin Zhao, Pratees…☆22Jun 19, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆14Jul 13, 2025Updated last year
- PSTensor provides a way to hack the memory management of tensors in TensorFlow and PyTorch by defining your own C++ Tensor Class.☆10Feb 10, 2022Updated 4 years ago
- A NCCL extension library, designed to efficiently offload GPU memory allocated by the NCCL communication library.☆119Dec 17, 2025Updated 9 months ago
- Implementation of a Tensorflow XLA rematerialization pass☆15Dec 20, 2019Updated 6 years ago
- Elastic Training on Kubernetes☆10Dec 13, 2021Updated 4 years ago
- Minimal, predictable, footgun-free config library.☆43Jul 12, 2026Updated 2 months ago
- Spatio-temporal pattern contruct and model fusion☆11Jun 10, 2019Updated 7 years ago
- Kaggleのshopeeコンペのリポジトリ☆11Jun 7, 2021Updated 5 years ago
- CMU 15-745 Spring 2014☆10Mar 7, 2014Updated 12 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- TairGis is a Redis Module that supports the query of intersection, contains, and within relationships between points, lines, and polygons…☆46Jun 5, 2025Updated last year
- ☆11Oct 8, 2022Updated 3 years ago
- APRIL: Active Partial Rollouts in Reinforcement Learning to Tame Long-tail Generation. A system-level optimization for scalable LLM tra…☆64Oct 11, 2025Updated 11 months ago
- ☆12Jul 4, 2024Updated 2 years ago
- ☆12Nov 5, 2024Updated last year
- MRT: Tracing the Evolution of Scientific Publications (TKDE 2021)☆18Mar 23, 2023Updated 3 years ago
- ☆13Aug 12, 2026Updated last month
- ☆19Jun 22, 2026Updated 2 months ago
- ☆18Nov 30, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆15Mar 12, 2023Updated 3 years ago
- A Controllable Model of Grounded Response Generation (AAAI 21)☆13Oct 25, 2022Updated 3 years ago
- A simple NUS Latex Beamer template.☆14Apr 15, 2024Updated 2 years ago
- implementation of dualformer☆25Mar 1, 2025Updated last year
- Turtlebot maze solver that senses environment through laser scans and navigates. A mini Project from Robot Ignite Academy Course☆10Aug 15, 2020Updated 6 years ago
- Extensions for the TG geometry library☆12Dec 3, 2024Updated last year
- SutroYaro — Sutro Group research workspace for energy-efficient AI training. Point any coding agent at the repo and it becomes a research…☆16May 29, 2026Updated 3 months ago
- TIER: Text-Image Encoder-based Regression for AIGC Image Quality Assessment☆10Mar 1, 2025Updated last year
- ☆21Jul 9, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code and data to explore neural scaling laws of xLSTM and Transformer models.☆24Apr 8, 2026Updated 5 months ago
- ☆97Sep 15, 2025Updated last year
- A Qwen .5B reasoning model trained on OpenR1-Math-220k☆14Updated this week
- Awesome-RL-Reasoning☆17Aug 26, 2026Updated 3 weeks ago
- CLAIR: A (surprisingly) simple semantic text metric with large language models.☆23Jan 28, 2024Updated 2 years ago
- An official PyTorch implementation of "Certifiably Robust Graph Contrastive Learning" (NeurIPS 2023)☆11Jan 22, 2024Updated 2 years ago
- Code of "Regularized Best-of-N Sampling with Minimum Bayes Risk Objective for Language Model Alignment" (2025).☆14Apr 4, 2025Updated last year