Fine-tunes a student LLM using teacher feedback for improved reasoning and answer quality. Implements GRPO with teacher-provided evaluations.
☆54May 7, 2025Updated last year
Alternatives and similar repositories for grpo-llm-evaluator
Users that are interested in grpo-llm-evaluator are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- NanoGPT (124M) in 5 minutes☆16Feb 14, 2025Updated last year
- this is based on the paper Chain-of-Retrieval Augmented Generation☆15Mar 29, 2025Updated last year
- Official implementation of PolySkill, a framework that enables web agents to learn generalizable and compositional skills through polymor…☆17Jul 6, 2026Updated 2 months ago
- ☆19Mar 10, 2025Updated last year
- ☆17Feb 1, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Apr 26, 2025Updated last year
- ☆13Mar 25, 2026Updated 5 months ago
- A simple one file python script that executes AI processes defined in YML.☆14Mar 26, 2023Updated 3 years ago
- This repository helps you evaluate your models on the FreshStack benchmark!☆34Dec 9, 2025Updated 8 months ago
- ☆15Jan 26, 2025Updated last year
- An easy-to-understand framework for LLM samplers that rewind and revise generated tokens☆151Jan 7, 2026Updated 8 months ago
- [NeurIPS 2024] Low rank memory efficient optimizer without SVD☆33Jul 1, 2025Updated last year
- Tensorflow tf.metrics tutorial☆12Aug 30, 2018Updated 8 years ago
- OpenCoconut implements a latent reasoning paradigm where we generate thoughts before decoding.☆173Jan 16, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Generating Easy-to-Understand Referring Expressions for Target Identifications☆18Aug 30, 2019Updated 7 years ago
- ☆24Jan 22, 2025Updated last year
- Tcurtsni: Reverse Instruction Chat, ever wonder what your LLM wants to ask you?☆23Jun 25, 2024Updated 2 years ago
- Learning adapter weights from task descriptions☆21Nov 12, 2023Updated 2 years ago
- chatGPT 'Autonomous Agent' in Node.js, written/runs in Termux. Sandboxed REPL access, Termux:API interface, chain-of-thought Question-Obs…☆16May 12, 2023Updated 3 years ago
- An introduction to LLM Sampling☆80Dec 15, 2024Updated last year
- Modify Entropy Based Sampling to work with Mac Silicon via MLX☆49Nov 6, 2024Updated last year
- Condense source code for LLM analysis by extracting essential highlights, utilizing a simplified version of Paul Gauthier's repomap techn…☆14Mar 3, 2024Updated 2 years ago
- All the content of my youtube channel : https://youtube.com/@florenzerstling?si=7t10PBr6MDha74PO☆14May 28, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆23May 25, 2023Updated 3 years ago
- ☆25Oct 10, 2025Updated 10 months ago
- [ACL 2025] Knowledge Unlearning for Large Language Models☆49Sep 18, 2025Updated 11 months ago
- A GitHub Template to generate AWS Serverless backends in Python☆10Jun 22, 2021Updated 5 years ago
- Exploring Applications of GRPO☆252Aug 25, 2025Updated last year
- minimal pytorch implementation of bm25 (with sparse tensors)☆105Oct 28, 2025Updated 10 months ago
- Implemention based on lightrag and nano-graphrag to connect with psql☆15Oct 28, 2024Updated last year
- Trading skills for Agentic AI☆15Jul 18, 2026Updated last month
- ☆16Mar 22, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆21Jul 25, 2025Updated last year
- Entropy Based Sampling and Parallel CoT Decoding☆17Oct 9, 2024Updated last year
- Repository of GUI Action Narrator☆13Apr 8, 2025Updated last year
- Explore training for quantized models☆28Jul 12, 2025Updated last year
- A collection of lightweight interpretability scripts to understand how LLMs think☆92Mar 18, 2026Updated 5 months ago
- MLX version of DINO DETR☆17Dec 26, 2024Updated last year
- ☆87Mar 23, 2026Updated 5 months ago