Tora: Torchtune-LoRA for RL
☆87Dec 2, 2025Updated 8 months ago
Alternatives and similar repositories for Tora
Users that are interested in Tora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Ludic – an LLM-RL library for the era of experience☆68Updated this week
- Project code for training LLMs to write better unit tests + code☆22May 19, 2025Updated last year
- General Reasoner: Advancing LLM Reasoning Across All Domains [NeurIPS25]☆229Nov 27, 2025Updated 8 months ago
- [ICLR 2026] Tina: Tiny Reasoning Models via LoRA☆338Sep 23, 2025Updated 10 months ago
- A comprehensive toolkit for streamlining data editing, search, and inspection for large-scale language model training and interpretabilit…☆21Oct 30, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code repository dedicated to experimenting and research with tiny reasoning language model☆52Nov 24, 2025Updated 8 months ago
- Github repo for MARVEL: Multidimensional Abstraction and Reasoning through Visual Evaluation and Learning☆18Jun 12, 2024Updated 2 years ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment☆18Dec 19, 2024Updated last year
- A curated list of resources on Reinforcement Learning with Verifiable Rewards (RLVR) and the reasoning capability boundary of Large Langu…☆91Dec 12, 2025Updated 7 months ago
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 4 months ago
- Synthetic pretraining data by rephrasing the web☆27Jun 5, 2026Updated 2 months ago
- ☆12Apr 10, 2024Updated 2 years ago
- ☆11Oct 25, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- UQ: Assessing Language Models on Unsolved Questions☆30Aug 26, 2025Updated 11 months ago
- ☆35Oct 23, 2025Updated 9 months ago
- This repo contains the source code for the paper "Evolution Strategies at Scale: LLM Fine-Tuning Beyond Reinforcement Learning"☆377Jun 26, 2026Updated last month
- Agentic RL Training at Scale☆1,878Updated this week
- rl from zero pretrain, can it be done? yes.☆295Sep 28, 2025Updated 10 months ago
- Minimal (truly) muP implementation, consistent with TP4 and TP5 papers notation☆14Jan 2, 2026Updated 7 months ago
- ☆19Jan 10, 2026Updated 7 months ago
- ☆75Jun 10, 2025Updated last year
- 🌾 OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.☆667Jan 29, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆65Mar 30, 2026Updated 4 months ago
- An efficent implementation of the method proposed in "The Era of 1-bit LLMs"☆155Oct 15, 2024Updated last year
- CodecHub: A Unified Library for Codec Models☆25Dec 24, 2025Updated 7 months ago
- ☆19Nov 4, 2025Updated 9 months ago
- The training codes of Jasper-Token-Compression-600M☆20Nov 19, 2025Updated 8 months ago
- "FiD-ICL: A Fusion-in-Decoder Approach for Efficient In-Context Learning" (ACL 2023)☆15Jul 24, 2023Updated 3 years ago
- Learning to route instances for Human vs AI Feedback (ACL Main '25)☆29Jul 23, 2025Updated last year
- ☆20Feb 2, 2026Updated 6 months ago
- Is In-Context Learning Sufficient for Instruction Following in LLMs? [ICLR 2025]☆33Jan 23, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆25Aug 20, 2025Updated 11 months ago
- ☆137Aug 19, 2024Updated last year
- A challenging aggregation benchmark for long-context models☆52Feb 22, 2026Updated 5 months ago
- MoE training for Me and You and maybe other people☆396Mar 15, 2026Updated 4 months ago
- Landing repository for the paper "Predicting the Order of Upcoming Tokens Improves Language Modeling"☆48May 13, 2026Updated 2 months ago
- [ICLR 2026]QeRL enables RL for 32B LLMs on a single H100 GPU.☆512Mar 30, 2026Updated 4 months ago
- Structured Primitives for Efficient Architecture Research☆21Dec 22, 2025Updated 7 months ago