☆46Dec 25, 2025Updated 9 months ago
Alternatives and similar repositories for retaining-by-doing
Users that are interested in retaining-by-doing are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆21Aug 25, 2026Updated last month
- Learning from Mixed Rollouts: Logit Fusion as a Bridge Between Imitation and Exploration☆18Feb 24, 2026Updated 7 months ago
- ☆16Oct 15, 2025Updated 11 months ago
- Rethinking the Trust Region in LLM Reinforcement Learning☆75Mar 2, 2026Updated 7 months ago
- Reasoning Activation in LLMs via Small Model Transfer (NeurIPS 2025)☆22Oct 16, 2025Updated 11 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- MOSS: Self-Evolution through Source-Level Rewriting in Autonomous Agent Systems☆24May 23, 2026Updated 4 months ago
- CoSCL: Cooperation of Small Continual Learners is Stronger than a Big One (ECCV 2022)☆20Sep 26, 2023Updated 3 years ago
- [ICLR25] Official Implementation of "Decoupling Angles and Strength in Low-rank Adaptation"☆15Dec 12, 2025Updated 9 months ago
- UFT: Unifying Supervised and Reinforcement Fine-Tuning☆34Jun 30, 2025Updated last year
- ☆17Oct 17, 2025Updated 11 months ago
- ☆29Aug 20, 2023Updated 3 years ago
- [NeurIPS 2024] Can Language Models Learn to Skip Steps?☆22Jan 25, 2025Updated last year
- Source code for "Gradient Based Memory Editing for Task-Free Continual Learning", 4th Lifelong ML Workshop@ICML 2020☆17Dec 8, 2022Updated 3 years ago
- (NBCE)Naive Bayes-based Context Extension on ChatGLM-6b☆15Jun 7, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Structuring Hour-Long Videos into Navigable Chapters and Hierarchical Summaries☆47Nov 19, 2025Updated 10 months ago
- [ICLR2026] DAG-Math: Graph-of-Thought Guided Mathematical Reasoning in LLMs☆24Oct 19, 2025Updated 11 months ago
- The implementation of ACL 2026 paper "Rethinking entropy interventions in rlvr: An entropy change perspective"☆27Jul 19, 2026Updated 2 months ago
- The official implementation of the CVPR'2024 work Interference-Free Low-Rank Adaptation for Continual Learning☆115Mar 13, 2025Updated last year
- Official implementation of 'RiskPO: Risk-based Policy Optimization via Verifiable Reward for LLM Post-Training', accepted by ICLR 2026☆19Oct 15, 2025Updated 11 months ago
- Code repo for paper: Effective Strategies for Asynchronous Software Engineering Agents☆75Apr 2, 2026Updated 6 months ago
- Source code for SWIFT, an efficient reward model.☆21Jan 13, 2026Updated 8 months ago
- Official [ICLR] Code Repository for "Gradient Projection Memory for Continual Learning"☆103Jun 25, 2021Updated 5 years ago
- Official PyTorch implementation of our CVPR 2025 paper, "LoRA Subtraction for Drift-Resistant Space in Exemplar-Free Continual Learning."☆19Mar 28, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- "Found in the Middle: How Language Models Use Long Contexts Better via Plug-and-Play Positional Encoding" Zhenyu Zhang, Runjin Chen, Shiw…☆35May 7, 2024Updated 2 years ago
- Human-centric environment representations from egocentric video☆15Feb 5, 2026Updated 8 months ago
- Paper list of compositional zero-shot learning☆11Jul 5, 2022Updated 4 years ago
- [CVPR 2025] CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answeri…☆62Jun 16, 2025Updated last year
- ☆33Feb 9, 2025Updated last year
- ☆17May 31, 2023Updated 3 years ago
- Codebase for the paper titled "Continual learning with local module selection"☆26Nov 15, 2021Updated 4 years ago
- DICE: Detecting In-distribution Data Contamination with LLM's Internal State☆11Sep 21, 2024Updated 2 years ago
- ☆12Sep 8, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official repository for Online Class Incremental Learning on Stochastic Blurry Task Boundary via Mask and Visual Prompt Tuning on ICCV 20…☆32Oct 26, 2024Updated last year
- [ACL 2024] The official codebase for the paper "Self-Distillation Bridges Distribution Gap in Language Model Fine-tuning".☆167Nov 2, 2024Updated last year
- Codebase for Linguistic Collapse: Neural Collapse in (Large) Language Models [NeurIPS 2024] [arXiv:2405.17767]☆18Apr 14, 2025Updated last year
- ☆11May 16, 2025Updated last year
- DNA-D2S: a systematic error simulation Model for DNA Data Storage channel☆12Feb 14, 2022Updated 4 years ago
- ☆101Jun 27, 2024Updated 2 years ago
- ☆14Apr 16, 2024Updated 2 years ago