[QReward] RewardService Python Client, make RL Training reward function more faster
☆16Apr 23, 2026Updated 4 months ago
Alternatives and similar repositories for QReward
Users that are interested in QReward are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A comprehensive medical AI evaluation framework based on GAPS methodology. Features automated assessment pipeline, thoracic surgery datas…☆46Nov 3, 2025Updated 9 months ago
- MrlX: A Multi-Agent Reinforcement Learning Framework☆221Jan 19, 2026Updated 7 months ago
- 国科大研究生课程 操作系统高级教程2023年思考题☆12Dec 24, 2023Updated 2 years ago
- My love.☆28Apr 1, 2026Updated 4 months ago
- [TBD] "m4: A Learned Flow-level Network Simulator" by Chenning Li, Anton A. Zabreyko, Om Chabra, Arash Nasr-Esfahany, Kevin Zhao, Pratees…☆21Jun 19, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A NCCL extension library, designed to efficiently offload GPU memory allocated by the NCCL communication library.☆116Dec 17, 2025Updated 8 months ago
- PSTensor provides a way to hack the memory management of tensors in TensorFlow and PyTorch by defining your own C++ Tensor Class.☆10Feb 10, 2022Updated 4 years ago
- Implementation of a Tensorflow XLA rematerialization pass☆15Dec 20, 2019Updated 6 years ago
- Elastic Training on Kubernetes☆10Dec 13, 2021Updated 4 years ago
- Minimal, predictable, footgun-free config library.☆43Jul 12, 2026Updated last month
- Spatio-temporal pattern contruct and model fusion☆11Jun 10, 2019Updated 7 years ago
- CMU 15-745 Spring 2014☆10Mar 7, 2014Updated 12 years ago
- AES-based-on-FPGA developed by verilog.☆22Apr 23, 2020Updated 6 years ago
- APRIL: Active Partial Rollouts in Reinforcement Learning to Tame Long-tail Generation. A system-level optimization for scalable LLM tra…☆63Oct 11, 2025Updated 10 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆12Jul 4, 2024Updated 2 years ago
- ☆15Jun 12, 2024Updated 2 years ago
- MRT: Tracing the Evolution of Scientific Publications (TKDE 2021)☆18Mar 23, 2023Updated 3 years ago
- ☆13Aug 12, 2026Updated 2 weeks ago
- ☆19Jun 22, 2026Updated 2 months ago
- ☆18Nov 30, 2025Updated 9 months ago
- MedResearcher-R1 is a deep research agent for medical scenarios, built on a knowledge-informed trajectory synthesis framework.☆520Sep 1, 2025Updated 11 months ago
- 让GIF显示文件自身的MD5值☆13Nov 22, 2020Updated 5 years ago
- ☆15Mar 12, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A Controllable Model of Grounded Response Generation (AAAI 21)☆13Oct 25, 2022Updated 3 years ago
- Turtlebot maze solver that senses environment through laser scans and navigates. A mini Project from Robot Ignite Academy Course☆10Aug 15, 2020Updated 6 years ago
- Extensions for the TG geometry library☆12Dec 3, 2024Updated last year
- TIER: Text-Image Encoder-based Regression for AIGC Image Quality Assessment☆10Mar 1, 2025Updated last year
- ☆12Updated this week
- ☆17Nov 15, 2025Updated 9 months ago
- Code and data to explore neural scaling laws of xLSTM and Transformer models.☆24Apr 8, 2026Updated 4 months ago
- ☆99Sep 15, 2025Updated 11 months ago
- A simple, stdlib only, Go module for generating RFC9562 UUIDs☆15Jul 28, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Awesome-RL-Reasoning☆17Updated this week
- An official PyTorch implementation of "Certifiably Robust Graph Contrastive Learning" (NeurIPS 2023)☆11Jan 22, 2024Updated 2 years ago
- ☆24May 22, 2024Updated 2 years ago
- Code of "Regularized Best-of-N Sampling with Minimum Bayes Risk Objective for Language Model Alignment" (2025).☆14Apr 4, 2025Updated last year
- ☆20Jul 12, 2023Updated 3 years ago
- [ICLR 2025] Drop-Upcycling: Training Sparse Mixture of Experts with Partial Re-initialization☆25Oct 5, 2025Updated 10 months ago
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆20Oct 23, 2023Updated 2 years ago