☆44Dec 25, 2025Updated 7 months ago
Alternatives and similar repositories for retaining-by-doing
Users that are interested in retaining-by-doing are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆21Mar 22, 2026Updated 4 months ago
- Learning from Mixed Rollouts: Logit Fusion as a Bridge Between Imitation and Exploration☆17Feb 24, 2026Updated 5 months ago
- [ICLR 2025] Unintentional Unalignment: Likelihood Displacement in Direct Preference Optimization☆32Jan 7, 2026Updated 7 months ago
- ☆15Oct 15, 2025Updated 9 months ago
- ☆16Nov 12, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Rethinking the Trust Region in LLM Reinforcement Learning☆67Mar 2, 2026Updated 5 months ago
- Reasoning Activation in LLMs via Small Model Transfer (NeurIPS 2025)☆22Oct 16, 2025Updated 9 months ago
- CoSCL: Cooperation of Small Continual Learners is Stronger than a Big One (ECCV 2022)☆20Sep 26, 2023Updated 2 years ago
- [ICLR25] Official Implementation of "Decoupling Angles and Strength in Low-rank Adaptation"☆15Dec 12, 2025Updated 7 months ago
- UFT: Unifying Supervised and Reinforcement Fine-Tuning☆31Jun 30, 2025Updated last year
- ☆29Aug 20, 2023Updated 2 years ago
- [NeurIPS 2024] Can Language Models Learn to Skip Steps?☆21Jan 25, 2025Updated last year
- Source code for "Gradient Based Memory Editing for Task-Free Continual Learning", 4th Lifelong ML Workshop@ICML 2020☆17Dec 8, 2022Updated 3 years ago
- Continual Memorization of Factoids in Large Language Models☆12Nov 20, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- (NBCE)Naive Bayes-based Context Extension on ChatGLM-6b☆15Jun 7, 2023Updated 3 years ago
- [ICLR2026] DAG-Math: Graph-of-Thought Guided Mathematical Reasoning in LLMs☆23Oct 19, 2025Updated 9 months ago
- The implementation of ACL 2026 paper "Rethinking entropy interventions in rlvr: An entropy change perspective"☆26Jul 19, 2026Updated 3 weeks ago
- The official implementation of the CVPR'2024 work Interference-Free Low-Rank Adaptation for Continual Learning☆113Mar 13, 2025Updated last year
- TOD-Flow: Modeling the Structure of Task-Oriented Dialogues☆13Feb 7, 2024Updated 2 years ago
- Bag of Instances Aggregation Boosts Self-supervised Distillation (ICLR 2022)☆33Apr 26, 2022Updated 4 years ago
- Official implementation of 'RiskPO: Risk-based Policy Optimization via Verifiable Reward for LLM Post-Training', accepted by ICLR 2026☆18Oct 15, 2025Updated 9 months ago
- Source code for SWIFT, an efficient reward model.☆21Jan 13, 2026Updated 6 months ago
- Official [ICLR] Code Repository for "Gradient Projection Memory for Continual Learning"☆102Jun 25, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- "Found in the Middle: How Language Models Use Long Contexts Better via Plug-and-Play Positional Encoding" Zhenyu Zhang, Runjin Chen, Shiw…☆35May 7, 2024Updated 2 years ago
- Human-centric environment representations from egocentric video☆15Feb 5, 2026Updated 6 months ago
- [ECCV2024] PyTorch implementation of "PILoRA: Prototype Guided Incremental LoRA for Federated Class-Incremental Learning"☆25Sep 25, 2024Updated last year
- [CVPR 2025] CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answeri…☆61Jun 16, 2025Updated last year
- ☆25May 8, 2025Updated last year
- ☆17May 31, 2023Updated 3 years ago
- Codebase for the paper titled "Continual learning with local module selection"☆26Nov 15, 2021Updated 4 years ago
- MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs☆17Jul 6, 2025Updated last year
- ☆43Jun 28, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆12Sep 8, 2023Updated 2 years ago
- Official repository for Online Class Incremental Learning on Stochastic Blurry Task Boundary via Mask and Visual Prompt Tuning on ICCV 20…☆32Oct 26, 2024Updated last year
- This is the official implementation of our NeurIPS 2025 paper "Gated Integration of Low-Rank Adaptation for Continual Learning of Large L…☆25Nov 27, 2025Updated 8 months ago
- CMU RavenClaw对话管理☆12Dec 13, 2017Updated 8 years ago
- The official implementation of NeurIPS2024 paper "SubgDiff: A Subgraph Diffusion Model to Improve Molecular Representation Learning."☆11May 28, 2025Updated last year
- [ACL 2024] The official codebase for the paper "Self-Distillation Bridges Distribution Gap in Language Model Fine-tuning".☆167Nov 2, 2024Updated last year
- Codebase for Linguistic Collapse: Neural Collapse in (Large) Language Models [NeurIPS 2024] [arXiv:2405.17767]☆18Apr 14, 2025Updated last year