Verifiers for LLM Reinforcement Learning
☆83Sep 11, 2025Updated 11 months ago
Alternatives and similar repositories for verifiers-deepresearch
Users that are interested in verifiers-deepresearch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Feb 23, 2026Updated 5 months ago
- ☆69May 23, 2025Updated last year
- Waffer-thin FlaskGPT on Vercel.☆12Jun 1, 2023Updated 3 years ago
- Perf monitoring CLI tool for Apple Silicon☆16Jan 1, 2024Updated 2 years ago
- ☆31Nov 14, 2025Updated 9 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Exploring Applications of GRPO☆252Aug 25, 2025Updated 11 months ago
- ☆21Mar 25, 2025Updated last year
- Official repo for "Binary Retrieval-augmented Reward Mitigates Hallucinations"☆16Nov 13, 2025Updated 9 months ago
- Improving Steering Vectors by Targeting Sparse Autoencoder Features☆29Nov 20, 2024Updated last year
- RAG Tool using Haystack, Mistral, and Chainlit. All open source stack on CPU.☆22Oct 14, 2023Updated 2 years ago
- rl from zero pretrain, can it be done? yes.☆295Sep 28, 2025Updated 10 months ago
- Build datasets using natural language☆589Sep 19, 2025Updated 10 months ago
- Seamless Voice Interactions with LLMs☆12Oct 28, 2023Updated 2 years ago
- Codebase exploration with AI research agents☆21Feb 25, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Project of ACL 2025 "UAlign: Leveraging Uncertainty Estimations for Factuality Alignment on Large Language Models"☆15Mar 25, 2025Updated last year
- ☆17Jan 30, 2024Updated 2 years ago
- ☆85Feb 1, 2024Updated 2 years ago
- The official repo of VideoAgentTrek☆58Oct 24, 2025Updated 9 months ago
- ☆21Nov 28, 2024Updated last year
- finetune your florence2 model easy☆21Jul 27, 2024Updated 2 years ago
- ☆26Nov 26, 2024Updated last year
- ☆28May 19, 2025Updated last year
- Testing paligemma2 finetuning on reasoning dataset☆18Dec 28, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.☆106Jul 19, 2025Updated last year
- macOS computer use CLI — screenshots, input simulation, app management, session orchestration☆24Mar 24, 2026Updated 4 months ago
- General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.☆17Apr 26, 2026Updated 3 months ago
- Visual demo of DSPy's prompt optimization on Gradio☆17Apr 14, 2025Updated last year
- Demo of knowledge graph creation and Graph RAG with Dspy and Kuzu☆22Jun 30, 2025Updated last year
- ☆19Jan 19, 2026Updated 6 months ago
- MCP server that creates its own tools as needed☆24Dec 26, 2024Updated last year
- Embedding models from Jina AI☆66Jan 18, 2024Updated 2 years ago
- ☆30Oct 7, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Redwood Research's transformer interpretability tools☆15Apr 15, 2022Updated 4 years ago
- Fused Qwen3 MoE layer for faster training, compatible with Transformers, LoRA, bnb 4-bit quant, Unsloth. Also possible to train LoRA over…☆258Jul 24, 2026Updated 3 weeks ago
- Implementing cognitive architecture and psychological memory concepts into Agentic LLM Systems☆550Dec 12, 2024Updated last year
- Implementation of Recursive Language Model paper from scratch☆46Feb 10, 2026Updated 6 months ago
- Evaluation kit for testing stateful agents☆83Updated this week
- A command-line utility to manage MLX models between your Hugging Face cache and LM Studio.☆88Nov 11, 2025Updated 9 months ago
- Luth is a state-of-the-art series of fine-tuned LLMs for French☆47Oct 12, 2025Updated 10 months ago