[NLPCC 2022] Kformer: Knowledge Injection in Transformer Feed-Forward Layers
☆39Oct 20, 2022Updated 3 years ago
Alternatives and similar repositories for Kformer
Users that are interested in Kformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Nov 15, 2022Updated 3 years ago
- Complexity Based Prompting for Multi-Step Reasoning☆17Mar 10, 2023Updated 3 years ago
- [ACL2023] Source code for Decouple knowledge from paramters for plug-and-play language modeling☆20Sep 18, 2023Updated 2 years ago
- Code and data repository for two papers (ACL & EMNLP 2024) on the topic of collapse in model editing.☆10Dec 20, 2024Updated last year
- Code for the ACL-2022 paper "Knowledge Neurons in Pretrained Transformers"☆177May 4, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Repository for "Propagating Knowledge Updates to LMs Through Distillation" (NeurIPS 2023).☆27Aug 25, 2024Updated 2 years ago
- FL-Tuning☆12Jul 11, 2022Updated 4 years ago
- The 4th rank system of the SemEval 2021 Task4.☆10May 7, 2022Updated 4 years ago
- MEND: Fast Model Editing at Scale☆260Aug 30, 2023Updated 3 years ago
- [AAAI 2026] ReCode: Reinforced Code Knowledge Editing for API Updates☆25Jul 1, 2025Updated last year
- ☆32Oct 17, 2022Updated 3 years ago
- Collection of Reverse Engineering in Large Model☆35Jan 8, 2025Updated last year
- BERT-based Biomedical Text Summarizer☆25Oct 2, 2019Updated 6 years ago
- Code for EMNLP2021 paper "Allocating Large Vocabulary Capacity for Cross-lingual Language Model Pre-training"☆20Nov 12, 2021Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration☆15Jun 4, 2024Updated 2 years ago
- [SIGIR 2023] Schema-aware Reference as Prompt Improves Data-Efficient Knowledge Graph Construction☆42Apr 5, 2023Updated 3 years ago
- This repository includes code for the paper "Does Localization Inform Editing? Surprising Differences in Where Knowledge Is Stored vs. Ca…☆62May 9, 2023Updated 3 years ago
- The official repository for our paper "The Dual Form of Neural Networks Revisited: Connecting Test Time Predictions to Training Patterns …☆16Jun 11, 2025Updated last year
- [NLPCC 2024] Shared Task 10: Regulating Large Language Models☆14Jun 12, 2024Updated 2 years ago
- [ACL 2023] Plug-and-Play Knowledge Injection for Pre-trained Language Models☆60Apr 1, 2024Updated 2 years ago
- ☆49Jan 21, 2024Updated 2 years ago
- Data and code for the SciFact-Open task☆29Nov 24, 2023Updated 2 years ago
- Claude Code 源码深度研究,包括 Foundations/Execution/Infrastructure 三大章节和 23 个子系统的架构分析拆解。☆77May 27, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆12Jan 25, 2024Updated 2 years ago
- NaturalProver: Grounded Mathematical Proof Generation with Language Models☆40Mar 24, 2023Updated 3 years ago
- [NLPCC 2023] Reasoning Through Memorization: Nearest Neighbor Knowledge Graph Embeddings with Language Models☆55Jul 31, 2023Updated 3 years ago
- ☆146Jul 21, 2024Updated 2 years ago
- [AAAI 2025] Assessing the Creativity of LLMs in Proposing Novel Solutions to Mathematical Problems☆13May 5, 2025Updated last year
- ☆36Jun 13, 2025Updated last year
- [NAACL'25] RuleR: Improving LLM Controllability by Rule-based Data Recycling☆14Sep 27, 2025Updated 11 months ago
- ☆175Jul 24, 2024Updated 2 years ago
- Context-based Dialogue Act Recognition using Recurrent Neural Networks☆13Nov 13, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [NeurIPS 2025 D&B Track] MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning Research☆36May 8, 2026Updated 3 months ago
- Code for "Unsupervised Enrichment of Persona-grounded Dialog with Background Stories", ACL 2021☆10Jul 8, 2021Updated 5 years ago
- This repository contains the data and code for the paper "SideControl: Controlled Open-domain Dialogue Generation via Additive Side Netwo…☆12Dec 1, 2021Updated 4 years ago
- ☆24Jun 7, 2021Updated 5 years ago
- ☆67May 18, 2023Updated 3 years ago
- [EMNLP 2023] Knowledge Rumination for Pre-trained Language Models☆17Jun 29, 2023Updated 3 years ago
- code for EMNLP 2024 paper: Neuron-Level Knowledge Attribution in Large Language Models☆52Nov 17, 2024Updated last year