RLVR Testing and Training
☆21Aug 28, 2025Updated 11 months ago
Alternatives and similar repositories for Reinforcement-learning-with-verifable-rewards-Learnings
Users that are interested in Reinforcement-learning-with-verifable-rewards-Learnings are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Evaluation framework for GGUF☆15Apr 2, 2026Updated 3 months ago
- Enhancing LLMs with LoRA☆224Oct 20, 2025Updated 9 months ago
- ☆15Mar 21, 2024Updated 2 years ago
- Convert documents, images to high-quality Markdown using Vision LLMs. Built for RAG ingestion pipelines.☆21Dec 20, 2025Updated 7 months ago
- High-Performance Text Deduplication Toolkit☆61Aug 25, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆22Jan 22, 2026Updated 6 months ago
- Evaluating practical performance of local multi-turn conversational LLMs.☆19Aug 1, 2025Updated 11 months ago
- Demo of building and intergraition MCP Server☆21Apr 9, 2025Updated last year
- highly structured NixOS configuration database☆15Feb 2, 2021Updated 5 years ago
- ☆11Oct 19, 2020Updated 5 years ago
- generate informative knowledge graph from text using open source models , ollama☆23Sep 1, 2025Updated 10 months ago
- The BAZAAR challenges LLMs to navigate the double-auction marketplace, where buyers and sellers must make strategic decisions with incomp…☆37Jul 30, 2025Updated 11 months ago
- Group-relative Trajectory-based Policy Optimization: Increasing Quality and Training Stability☆42Feb 23, 2026Updated 5 months ago
- C++ version of pyannote audio overlapped speech detection pipeline☆13Feb 14, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- An agent for playing Atari games running on a Teensy microcontroller☆14Nov 11, 2022Updated 3 years ago
- A browser extension to automatically extract and batch-copy transcripts from Udemy videos. Perfect for students and learners who use note…☆15Jun 14, 2026Updated last month
- make some nice fire in webgpu☆17Apr 26, 2026Updated 3 months ago
- Multi-vector latent space steering adapter module for language models☆20Nov 22, 2025Updated 8 months ago
- Lightweight C inference for Qwen3 GGUF. Multiturn prefix caching & batch processing.☆25Sep 1, 2025Updated 10 months ago
- Mass-edit files with LLMs☆26Jun 2, 2026Updated last month
- Llama cute voice assistant☆28Sep 10, 2023Updated 2 years ago
- A Multi-Agentic AI Assistant/Builder☆28Jun 1, 2026Updated last month
- Simple and fast server for GPTQ-quantized LLaMA inference☆24May 18, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆12Mar 25, 2025Updated last year
- An API for VoiceCraft.☆25Jun 27, 2024Updated 2 years ago
- ☆26Sep 4, 2025Updated 10 months ago
- Mini RL Lab☆16Jun 17, 2024Updated 2 years ago
- Experience the power of AI with this free AI voice generator demo. Utilizing Deepgram and Groq, we transform text into voice seamlessly. …☆38Jun 12, 2024Updated 2 years ago
- Zero-instrumentation LLM API and MCP tracer for your agents powered by eBPF — latency, tokens, and tool use in realtime☆18Mar 16, 2026Updated 4 months ago
- Depth Estimation model, DepthPro by Apple, trained for Image Segmentation and Image Super Resolution.☆16Jan 2, 2025Updated last year
- Create topological graph for image segments.☆23Sep 28, 2024Updated last year
- Raspberry Pi HUD/dashboard framework☆20Sep 15, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Vintix: Action Model via In-Context Reinforcement Learning - - — ICML 2025☆51May 23, 2025Updated last year
- LLM Context Manager for inference optimization☆25Jul 28, 2025Updated last year
- Zero-shot forecasting, tabular classification, and regression via MCP — exposes Google TimesFM 2.5 and TabFM v1.0.0 to AI assistants. Jus…☆25Jul 12, 2026Updated 2 weeks ago
- ☆39Nov 10, 2025Updated 8 months ago
- The official repository for the ARSENAL model☆16Jul 3, 2026Updated 3 weeks ago
- Geometric AI research: a proven cube-math core, reusable vector-collapse dynamics, and reproducible experiments in embeddings, NLI, gener…☆15Updated this week
- Host LLM via text-generation-inference☆16Dec 5, 2023Updated 2 years ago