Tora: Torchtune-LoRA for RL
☆87Dec 2, 2025Updated 7 months ago
Alternatives and similar repositories for Tora
Users that are interested in Tora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Ludic – an LLM-RL library for the era of experience☆67Jan 9, 2026Updated 6 months ago
- Project code for training LLMs to write better unit tests + code☆22May 19, 2025Updated last year
- General Reasoner: Advancing LLM Reasoning Across All Domains [NeurIPS25]☆228Nov 27, 2025Updated 7 months ago
- [ICLR 2026] Tina: Tiny Reasoning Models via LoRA☆338Sep 23, 2025Updated 9 months ago
- A comprehensive toolkit for streamlining data editing, search, and inspection for large-scale language model training and interpretabilit…☆21Oct 30, 2025Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code repository dedicated to experimenting and research with tiny reasoning language model☆52Nov 24, 2025Updated 7 months ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- A curated list of resources on Reinforcement Learning with Verifiable Rewards (RLVR) and the reasoning capability boundary of Large Langu…☆89Dec 12, 2025Updated 7 months ago
- RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment☆18Dec 19, 2024Updated last year
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 3 months ago
- Synthetic pretraining data by rephrasing the web☆24Jun 5, 2026Updated last month
- The official code implementation of the ACL2025 paper “A Text is Worth Several Tokens: Text Embedding from LLMs Secretly Aligns Well with…☆18Jul 12, 2025Updated last year
- Descript Audio Codec - VAE Variant (.dac-vae): High-Fidelity Audio Compression with Variational Autoencoder☆38Aug 30, 2025Updated 10 months ago
- UQ: Assessing Language Models on Unsolved Questions☆30Aug 26, 2025Updated 10 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆35Oct 23, 2025Updated 8 months ago
- This repo contains the source code for the paper "Evolution Strategies at Scale: LLM Fine-Tuning Beyond Reinforcement Learning"☆373Jun 26, 2026Updated 3 weeks ago
- rl from zero pretrain, can it be done? yes.☆295Sep 28, 2025Updated 9 months ago
- Minimal (truly) muP implementation, consistent with TP4 and TP5 papers notation☆14Jan 2, 2026Updated 6 months ago
- ☆76Jun 10, 2025Updated last year
- 🌾 OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.☆667Jan 29, 2026Updated 5 months ago
- ☆65Mar 30, 2026Updated 3 months ago
- An efficent implementation of the method proposed in "The Era of 1-bit LLMs"☆154Oct 15, 2024Updated last year
- Official PyTorch implementation and models for paper "Diffusion Beats Autoregressive in Data-Constrained Settings". We find diffusion mod…☆127Jan 10, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The training codes of Jasper-Token-Compression-600M☆20Nov 19, 2025Updated 8 months ago
- ☆19Nov 4, 2025Updated 8 months ago
- ☆20Feb 2, 2026Updated 5 months ago
- "FiD-ICL: A Fusion-in-Decoder Approach for Efficient In-Context Learning" (ACL 2023)☆15Jul 24, 2023Updated 2 years ago
- Learning to route instances for Human vs AI Feedback (ACL Main '25)☆29Jul 23, 2025Updated 11 months ago
- Is In-Context Learning Sufficient for Instruction Following in LLMs? [ICLR 2025]☆33Jan 23, 2025Updated last year
- Low memory full parameter finetuning of LLMs☆54Jul 18, 2025Updated last year
- ☆24Aug 20, 2025Updated 11 months ago
- ☆137Aug 19, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Source code for the paper "Positional Attention: Expressivity and Learnability of Algorithmic Computation"☆14May 26, 2025Updated last year
- A challenging aggregation benchmark for long-context models☆51Feb 22, 2026Updated 4 months ago
- Landing repository for the paper "Predicting the Order of Upcoming Tokens Improves Language Modeling"☆48May 13, 2026Updated 2 months ago
- MoE training for Me and You and maybe other people☆394Mar 15, 2026Updated 4 months ago
- A framework to meta-train transformers for causal ICL☆11Jul 15, 2026Updated last week
- [ICML 2026 Spotlight] On the Interplay of Pre-Training, Mid-Training, and RL on Reasoning Language Models☆162Jun 8, 2026Updated last month
- [ICLR 2026]QeRL enables RL for 32B LLMs on a single H100 GPU.☆511Mar 30, 2026Updated 3 months ago