Project code for training LLMs to write better unit tests + code
☆22May 19, 2025Updated last year
Alternatives and similar repositories for unit_test_rl
Users that are interested in unit_test_rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Apr 26, 2025Updated last year
- ☆20Oct 25, 2025Updated 10 months ago
- Optimizing Causal LMs through GRPO with weighted reward functions and automated hyperparameter tuning using Optuna☆60Oct 18, 2025Updated 11 months ago
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 5 months ago
- General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.☆17Apr 26, 2026Updated 4 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- MMLU-Pro eval results☆15Aug 21, 2025Updated last year
- The best benchmark for LLMs on Apple's MLX framework knowledge and coding tasks.☆38Jun 12, 2026Updated 3 months ago
- a set of scripts to easily convert all training data from huggingface into alpaca instruct or sharegpt format, which should allow for eas…☆20Mar 14, 2025Updated last year
- ☆69May 23, 2025Updated last year
- Examples for how to use DSPY and GEPA☆47Apr 6, 2026Updated 5 months ago
- A simple MLX implementation for pretraining LLMs on Apple Silicon.☆84Aug 20, 2025Updated last year
- Triton‑style kernel toolkit for MLX plus a small upstream incubator: prototype, benchmark, and upstream fusions for Apple Silicon☆53Mar 31, 2026Updated 5 months ago
- Implementation of MixCE method described in ACL 2023 paper by Zhang et al.☆20May 29, 2023Updated 3 years ago
- Recursive AI System☆52Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆27May 7, 2025Updated last year
- ☆30Oct 24, 2025Updated 10 months ago
- ☆19Mar 16, 2025Updated last year
- ☆29Jan 19, 2026Updated 8 months ago
- Use AI to edit your documents in real-time. Provide feedback and let the AI do all the work.☆29Jul 24, 2024Updated 2 years ago
- Exploring Applications of GRPO☆251Aug 25, 2025Updated last year
- ☆29Aug 27, 2025Updated last year
- Electron shell for mrmd - Zen Markdown Editor with real-time collaboration☆37Mar 20, 2026Updated 6 months ago
- Ultra low overhead NVIDIA GPU telemetry plugin for telegraf with memory temperature readings.☆63Jul 8, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Distil SN97 — Competitive Model Distillation on Bittensor☆35May 20, 2026Updated 4 months ago
- qwen3-base family of models RL on gsm8k using verl, is there an RL power law on downstream tasks?☆27Oct 19, 2025Updated 11 months ago
- ☆146Aug 20, 2025Updated last year
- A framework for optimizing DSPy programs with RL☆339Jan 12, 2026Updated 8 months ago
- Very minimal (and stateless) agent framework☆44Jan 12, 2025Updated last year
- Tora: Torchtune-LoRA for RL☆87Dec 2, 2025Updated 9 months ago
- PolarEngine: vLLM plugin for PolarQuant quantized LLM inference — 75% FP16 speed at 2.3x less VRAM☆36Apr 13, 2026Updated 5 months ago
- Create embeddings for LLM using the Nomic API☆23Nov 21, 2024Updated last year
- Distributed Reinforcement Learning for LLM Fine-Tuning with multi-GPU utilization☆22Mar 12, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Agentic workflow for tackling all open Erdos problems at once.☆34May 10, 2026Updated 4 months ago
- ☆10Oct 24, 2024Updated last year
- ☆13Jun 29, 2024Updated 2 years ago
- An experiment to see if chatgpt can improve the output of the stanford alpaca dataset☆12Mar 29, 2023Updated 3 years ago
- MLX Implementation of Recursive Reasoning with Tiny Networks☆79Oct 11, 2025Updated 11 months ago
- A sample app to debug and validate cellular modems on balena devices☆13Jun 5, 2019Updated 7 years ago
- ☆46Feb 20, 2026Updated 7 months ago