Tora: Torchtune-LoRA for RL
☆87Dec 2, 2025Updated 8 months ago
Alternatives and similar repositories for Tora
Users that are interested in Tora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Ludic – an LLM-RL library for the era of experience☆67Aug 9, 2026Updated 3 weeks ago
- Luth is a state-of-the-art series of fine-tuned LLMs for French☆47Oct 12, 2025Updated 10 months ago
- General Reasoner: Advancing LLM Reasoning Across All Domains [NeurIPS25]☆231Nov 27, 2025Updated 9 months ago
- [ICLR 2026] Tina: Tiny Reasoning Models via LoRA☆337Sep 23, 2025Updated 11 months ago
- A comprehensive toolkit for streamlining data editing, search, and inspection for large-scale language model training and interpretabilit…☆21Oct 30, 2025Updated 10 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Code repository dedicated to experimenting and research with tiny reasoning language model☆52Nov 24, 2025Updated 9 months ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment☆18Dec 19, 2024Updated last year
- A curated list of resources on Reinforcement Learning with Verifiable Rewards (RLVR) and the reasoning capability boundary of Large Langu…☆92Dec 12, 2025Updated 8 months ago
- ☆12Apr 10, 2024Updated 2 years ago
- UQ: Assessing Language Models on Unsolved Questions☆30Aug 26, 2025Updated last year
- ☆35Oct 23, 2025Updated 10 months ago
- This repo contains the source code for the paper "Evolution Strategies at Scale: LLM Fine-Tuning Beyond Reinforcement Learning"☆384Jun 26, 2026Updated 2 months ago
- Agentic RL Training at Scale☆1,993Updated this week
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- rl from zero pretrain, can it be done? yes.☆296Sep 28, 2025Updated 11 months ago
- Minimal (truly) muP implementation, consistent with TP4 and TP5 papers notation☆14Jan 2, 2026Updated 7 months ago
- ☆19Jan 10, 2026Updated 7 months ago
- ☆75Jun 10, 2025Updated last year
- 🌾 OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.☆670Jan 29, 2026Updated 7 months ago
- ☆65Mar 30, 2026Updated 5 months ago
- An efficent implementation of the method proposed in "The Era of 1-bit LLMs"☆155Oct 15, 2024Updated last year
- Official PyTorch implementation and models for paper "Diffusion Beats Autoregressive in Data-Constrained Settings". We find diffusion mod…☆128Jan 10, 2026Updated 7 months ago
- ☆20Nov 4, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The training codes of Jasper-Token-Compression-600M☆21Nov 19, 2025Updated 9 months ago
- "FiD-ICL: A Fusion-in-Decoder Approach for Efficient In-Context Learning" (ACL 2023)☆15Jul 24, 2023Updated 3 years ago
- Learning to route instances for Human vs AI Feedback (ACL Main '25)☆30Jul 23, 2025Updated last year
- ☆20Feb 2, 2026Updated 6 months ago
- Low memory full parameter finetuning of LLMs☆54Jul 18, 2025Updated last year
- The official repo for the paper "Teacher Forcing Recovers Reward Functions for Text Generation"☆31May 27, 2023Updated 3 years ago
- ☆136Aug 19, 2024Updated 2 years ago
- Source code for the paper "Positional Attention: Expressivity and Learnability of Algorithmic Computation"☆14May 26, 2025Updated last year
- Stress-Testing Image Generation Models with Explainable Human Evaluation on Open-ended Real-World Tasks [ICLR 2026]☆32Apr 2, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A challenging aggregation benchmark for long-context models☆53Feb 22, 2026Updated 6 months ago
- MoE training for Me and You and maybe other people☆396Mar 15, 2026Updated 5 months ago
- Landing repository for the paper "Predicting the Order of Upcoming Tokens Improves Language Modeling"☆48May 13, 2026Updated 3 months ago
- A framework to meta-train transformers for causal ICL☆11Jul 15, 2026Updated last month
- [ICLR 2026]QeRL enables RL for 32B LLMs on a single H100 GPU.☆516Mar 30, 2026Updated 5 months ago
- Structured Primitives for Efficient Architecture Research☆21Dec 22, 2025Updated 8 months ago
- H-Net Dynamic Hierarchical Architecture☆81Sep 11, 2025Updated 11 months ago