TRACE: A Comprehensive Benchmark for Continual Learning in Large Language Models
☆101Jan 24, 2024Updated 2 years ago
Alternatives and similar repositories for TRACE
Users that are interested in TRACE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆209Jul 13, 2024Updated 2 years ago
- Code for "Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized Rehearsal" (ACL 2024)☆17Oct 21, 2024Updated last year
- [ACL'25 Main] Official Implementation of HiDe-LLaVA: Hierarchical Decoupling for Continual Instruction Tuning of Multimodal Large Languag…☆55Jun 1, 2026Updated last month
- [ECCV 2024] MagMax: Leveraging Model Merging for Seamless Continual Learning (official repository)☆32Jul 29, 2024Updated last year
- This is the official code for Any-SSR "Analytic Subspace Routing: How Recursive Least Squares Works in Continual Learning of Large Langua…☆27Jan 31, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆11Nov 11, 2022Updated 3 years ago
- Code for ACL 2024 accepted paper titled "SAPT: A Shared Attention Framework for Parameter-Efficient Continual Learning of Large Language …☆40Jan 13, 2025Updated last year
- [ACL2024] A Codebase for Incremental Learning with Large Language Models; Official released code for "Learn or Recall? Revisiting Increme…☆62Feb 1, 2025Updated last year
- ☆10Feb 6, 2025Updated last year
- A mobile application that can help users get the perfect blackboard photos.☆25Jun 2, 2024Updated 2 years ago
- Toward Multi Modality Language Model - implementation of GPT-4o/Project Astra☆16Dec 10, 2024Updated last year
- [ICML 2023] Parameter-Level Soft-Masking for Continual Learning☆19Jul 13, 2023Updated 3 years ago
- Code to reproduce the experiments of "Rethinking Experience Replay: a Bag of Tricks for Continual Learning"☆54Feb 16, 2023Updated 3 years ago
- Instruction Tuning in Continual Learning paradigm☆80Feb 5, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Analyzing and Reducing Catastrophic Forgetting in Parameter Efficient Tuning☆38Nov 17, 2024Updated last year
- Adding new tasks to T0 without catastrophic forgetting☆33Oct 20, 2022Updated 3 years ago
- Official [AAAI] Code Repository for "Continual Learning with Scaled Gradient Projection".☆17Jun 28, 2023Updated 3 years ago
- Learn from Experience skill 是让 Agent 从经验中学习 -- 纠正过的错误不再重犯,好的方法自动沉淀,越用越懂你。原理是基于你跟Agent交互中的经验(踩坑、纠正、好方法)记下来、整理好、用起来,不再随会话结束而消失。☆20Apr 17, 2026Updated 3 months ago
- Official code for PLoP☆20Mar 6, 2026Updated 4 months ago
- Progressive Prompts: Continual Learning for Language Models☆96Apr 24, 2023Updated 3 years ago
- An unofficial implementation of "Mixture-of-Depths: Dynamically allocating compute in transformer-based language models"☆35Jun 7, 2024Updated 2 years ago
- MINER: Mutual Information based Named Entity Recognition☆37May 24, 2022Updated 4 years ago
- Source code for a LoRA-based continual relation extraction method.☆14Sep 25, 2023Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Hierarchical Decomposition of Prompt-Based Continual Learning: Rethinking Obscured Sub-optimality (NeurIPS 2023, Spotlight)☆92Nov 15, 2024Updated last year
- ☆16Jun 1, 2023Updated 3 years ago
- An Extendible (General) Continual Learning Framework based on Pytorch - official codebase of Dark Experience for General Continual Learni…☆825May 20, 2026Updated 2 months ago
- [CVPR 2025] CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answeri…☆60Jun 16, 2025Updated last year
- [AAAI2024] Summarizing Stream Data for Memory-Restricted Online Continual Learning☆22Apr 30, 2024Updated 2 years ago
- Official code for the paper CipherDAug: Ciphertext based Data Augmentation for Neural Machine Translation published at ACL 2022 main conf…☆12Apr 6, 2023Updated 3 years ago
- Service for Bert model to Vector. 高效的文本转向量(Text-To-Vector)服务,支持GPU多卡、多worker、多客户端调用,开箱即用。☆12May 24, 2022Updated 4 years ago
- ☆20Mar 12, 2025Updated last year
- 1.4B sLLM for Chinese and English - HammerLLM🔨☆44Apr 7, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- WebRED is a large and diverse manually annotated dataset for extracting relationships from a variety of text found on the World Wide Web.☆22Mar 11, 2021Updated 5 years ago
- [ICLR 2024 Oral] Improving Convergence and Generalization Using Parameter Symmetries☆31May 29, 2024Updated 2 years ago
- Implementation of the article Online Structured Laplace Approximations For Overcoming Catastrophic Forgetting (https://arxiv.org/pdf/1805…☆10Feb 1, 2019Updated 7 years ago
- A tool for extracting plain text and internal Wikipedia links from Wikipedia dumps☆11Apr 18, 2019Updated 7 years ago
- [ICLR 2025] A Closer Look at Machine Unlearning for Large Language Models☆49Dec 4, 2024Updated last year
- Bayesian low-rank adaptation for large language models☆29May 4, 2024Updated 2 years ago
- Code for the paper "Modelling Latent Translations for Cross-Lingual Transfer"☆17Nov 22, 2021Updated 4 years ago