RLVR Testing and Training
☆21Aug 28, 2025Updated last year
Alternatives and similar repositories for Reinforcement-learning-with-verifable-rewards-Learnings
Users that are interested in Reinforcement-learning-with-verifable-rewards-Learnings are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Enhancing LLMs with LoRA☆227Oct 20, 2025Updated 11 months ago
- ☆15Mar 21, 2024Updated 2 years ago
- Convert documents, images to high-quality Markdown using Vision LLMs. Built for RAG ingestion pipelines.☆21Updated this week
- MiRAGE: A Multiagent Framework for Generating Multimodal Multihop Question-Answer Dataset for RAG Evaluation☆24Aug 5, 2026Updated last month
- High-Performance Text Deduplication Toolkit☆60Aug 25, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A novel hybrid AI architecture leveraging Titan's-like memory and HRM-like reasoning☆27Aug 28, 2026Updated 3 weeks ago
- ☆22Jan 22, 2026Updated 8 months ago
- Evaluating practical performance of local multi-turn conversational LLMs.☆20Aug 1, 2025Updated last year
- ☆11Oct 19, 2020Updated 5 years ago
- ☆14May 25, 2023Updated 3 years ago
- generate informative knowledge graph from text using open source models , ollama☆23Sep 1, 2025Updated last year
- The BAZAAR challenges LLMs to navigate the double-auction marketplace, where buyers and sellers must make strategic decisions with incomp…☆39Jul 30, 2025Updated last year
- An experimental desktop client for using Claude Desktop's MCP with Novelcrafter codices.☆11Dec 3, 2024Updated last year
- Official code for "A General Learning Framework for Open Ad Hoc Teamwork Using Graph-based Policy Learning"☆15Mar 1, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- make some nice fire in webgpu☆17Apr 26, 2026Updated 5 months ago
- SiDeGame - Simplified Defusal Game☆12Apr 17, 2025Updated last year
- An agent for playing Atari games running on a Teensy microcontroller☆14Nov 11, 2022Updated 3 years ago
- Open source code combining implementations of Upside Down Reinforcement Learning and Reward Conditioned Policies☆19Mar 10, 2021Updated 5 years ago
- Multi-vector latent space steering adapter module for language models☆20Nov 22, 2025Updated 10 months ago
- ☆16Sep 28, 2023Updated 2 years ago
- A Multi-Agentic AI Assistant/Builder☆28Jun 1, 2026Updated 3 months ago
- ☆13Aug 13, 2025Updated last year
- Create and improve AI agents that get better over time with automatic optimization and continuous learning☆43Updated this week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆24Sep 20, 2025Updated last year
- Simple and fast server for GPTQ-quantized LLaMA inference☆24May 18, 2023Updated 3 years ago
- ☆13Mar 25, 2025Updated last year
- Modelling heterogeneous distributions with an Uncountable Mixture of Asymmetric Laplacians☆18Oct 27, 2019Updated 6 years ago
- An API for VoiceCraft.☆25Jun 27, 2024Updated 2 years ago
- An open source deep learning library for Unity.☆16Sep 13, 2026Updated 2 weeks ago
- ☆32Sep 4, 2025Updated last year
- Mini RL Lab☆16Jun 17, 2024Updated 2 years ago
- Experience the power of AI with this free AI voice generator demo. Utilizing Deepgram and Groq, we transform text into voice seamlessly. …☆38Jun 12, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Graph convolutional memory☆16May 26, 2022Updated 4 years ago
- Zero-instrumentation LLM API and MCP tracer for your agents powered by eBPF — latency, tokens, and tool use in realtime☆18Mar 16, 2026Updated 6 months ago
- Fulloch - The Fully Local Home Voice Assistant☆157Updated this week
- Vintix: Action Model via In-Context Reinforcement Learning - - — ICML 2025☆52May 23, 2025Updated last year
- LLM Context Manager for inference optimization☆25Jul 28, 2025Updated last year
- utilizing RL and GNN for trajectory planning(co-work)☆13Jul 28, 2023Updated 3 years ago
- A self-hosted todo-list web app for personal daily use, inspired by Things for macOS and currently in development.☆17Jul 6, 2026Updated 2 months ago