☆34Aug 5, 2023Updated 3 years ago
Alternatives and similar repositories for Transformer-Patcher
Users that are interested in Transformer-Patcher are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Dataset for Unified Editing, EMNLP 2023. This is a model editing dataset where edits are natural language phrases.☆24Sep 4, 2024Updated last year
- Mass-editing thousands of facts into a transformer memory (ICLR 2023)☆557Jan 31, 2024Updated 2 years ago
- ☆14Feb 12, 2024Updated 2 years ago
- ☆38Jan 26, 2024Updated 2 years ago
- Semi-Parametric Editing with a Retrieval-Augmented Counterfactual Model☆72Nov 1, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆18Mar 3, 2025Updated last year
- [Findings of EMNLP 2022] Code of paper Generative Prompt Tuning for Relation Classification. https://arxiv.org/abs/2210.12435☆20May 7, 2023Updated 3 years ago
- ☆17Nov 7, 2023Updated 2 years ago
- ☆17Aug 2, 2023Updated 3 years ago
- [NeurIPS'23] Aging with GRACE: Lifelong Model Editing with Discrete Key-Value Adaptors☆86Dec 21, 2024Updated last year
- [AAAI 2025 oral] Attribution Analysis Meets Model Editing: Advancing Knowledge Correction in Vision Language Models with VisEdit☆19Apr 19, 2025Updated last year
- Repository for "Propagating Knowledge Updates to LMs Through Distillation" (NeurIPS 2023).☆27Aug 25, 2024Updated last year
- Evaluating the Ripple Effects of Knowledge Editing in Language Models☆57Apr 15, 2024Updated 2 years ago
- [AAAI 2024] MELO: Enhancing Model Editing with Neuron-indexed Dynamic LoRA☆28Apr 9, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆14Nov 15, 2022Updated 3 years ago
- Data creation, training and eval scripts for the IRCoder paper☆21May 31, 2024Updated 2 years ago
- [EMNLP 2023] MQuAKE: Assessing Knowledge Editing in Language Models via Multi-Hop Questions☆124Sep 12, 2024Updated last year
- ☆11Jun 16, 2024Updated 2 years ago
- ☆29Jul 16, 2024Updated 2 years ago
- [ACL 2023 Findings] What In-Context Learning “Learns” In-Context: Disentangling Task Recognition and Task Learning☆21Jul 9, 2023Updated 3 years ago
- [NLPCC 2022] Kformer: Knowledge Injection in Transformer Feed-Forward Layers☆39Oct 20, 2022Updated 3 years ago
- A controlled benchmark on evaluating and studying the dynamics of Long Context Language Models☆26Oct 17, 2025Updated 9 months ago
- Paper list of "The Life Cycle of Knowledge in Big Language Models: A Survey"☆58Aug 24, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The official repository for our paper "The Dual Form of Neural Networks Revisited: Connecting Test Time Predictions to Training Patterns …☆16Jun 11, 2025Updated last year
- Interpretable unified language safety checking with large language models☆32Apr 15, 2023Updated 3 years ago
- [NLPCC 2024] Shared Task 10: Regulating Large Language Models☆14Jun 12, 2024Updated 2 years ago
- ParetoDrug☆11Sep 3, 2024Updated last year
- ☆41Nov 30, 2023Updated 2 years ago
- ☆20May 30, 2024Updated 2 years ago
- Third Person Shooter for Unity☆13Jun 26, 2022Updated 4 years ago
- ☆32Oct 17, 2022Updated 3 years ago
- PathPiece tokenizer☆14Nov 10, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆36Jun 13, 2025Updated last year
- This repository provides the code for applying Contrastive Learning Penalty Loss (CLPL) and Mixture of Experts (MoE) to the BGE-M3 text e…☆11Dec 27, 2024Updated last year
- Must-read Papers on Knowledge Editing for Large Language Models.☆1,246Jun 25, 2026Updated last month
- This is the paddle code for SeBoW(Self-Born wiring for neural trees), a kind of neural tree born form a large search space☆11Dec 10, 2021Updated 4 years ago
- Source code repo for paper "TLDR: Token Loss Dynamic Reweighting for Reducing Repetitive Utterance Generation"☆10Aug 11, 2023Updated 3 years ago
- [ACL 2023] Contextual Distortion Reveals Constituency: Mask Language Models are Implicit Parsers.☆14Jun 3, 2023Updated 3 years ago
- ☆19Dec 12, 2025Updated 8 months ago