[ACL 2025] How Do LLMs Acquire New Knowledge? A Knowledge Circuits Perspective on Continual Pre-Training
☆50Jul 18, 2025Updated last year
Alternatives and similar repositories for DynamicKnowledgeCircuits
Users that are interested in DynamicKnowledgeCircuits are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2025] Circuit-Aware Editing Enables Generalizable Knowledge Learners☆19Nov 17, 2025Updated 8 months ago
- [NeurIPS 2024] Knowledge Circuits in Pretrained Transformers☆172Nov 14, 2025Updated 8 months ago
- Code repository for the paper "The Inherent Limits of Pretrained LLMs: The Unexpected Convergence of Instruction Tuning and In-Context Le…☆14Jan 16, 2025Updated last year
- [ICLR 2024] Unveiling the Pitfalls of Knowledge Editing for Large Language Models☆22Jun 13, 2024Updated 2 years ago
- ☆29Feb 27, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The 4th rank system of the SemEval 2021 Task4.☆10May 7, 2022Updated 4 years ago
- [WWW 2026 Oral] MoE-CL:Self-Evolving LLMs via Continual Instruction Tuning☆21Dec 1, 2025Updated 7 months ago
- ☆19May 17, 2025Updated last year
- Evaluating the Ripple Effects of Knowledge Editing in Language Models☆57Apr 15, 2024Updated 2 years ago
- [NeurIPS'23] Aging with GRACE: Lifelong Model Editing with Discrete Key-Value Adaptors☆86Dec 21, 2024Updated last year
- Schoenfeld’s Anatomy of Mathematical Reasoning by Language Models☆27Dec 21, 2025Updated 7 months ago
- A Mechanistic‑Interpretability study that finds the structural dynamics of Large Language Models under fine‑tuning.☆17May 30, 2025Updated last year
- ☆82May 23, 2026Updated last month
- This repository contains the resource introduced in the paper: "Truth or Mirage? Towards End-to-End Factuality Evaluation with LLM-Oasis"…☆25Oct 15, 2025Updated 9 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Advanced SQLMap command builder with an intuitive cheatsheet UI. Works locally in your browser as a single HTML file (no data sent anywhe…☆40Jul 6, 2025Updated last year
- CogKTR: A Knowledge-Enhanced Text Representation Toolkit for Natural Language Understanding. EMNLP 2022☆32Oct 14, 2022Updated 3 years ago
- ☆18Dec 23, 2025Updated 6 months ago
- ☆33Oct 2, 2025Updated 9 months ago
- TLS & API keys for your LLM APIs☆20Dec 17, 2025Updated 7 months ago
- ☆19Mar 10, 2025Updated last year
- Flax (JAX) implementation of Progressive Growing of GANs for Improved Quality, Stability, and Variation☆12May 24, 2021Updated 5 years ago
- Recipes to train the self-rewarding reasoning LLMs.☆231Mar 2, 2025Updated last year
- Aligning Agentic World Models via Knowledgeable Experience Learning☆37May 15, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆12Jul 3, 2023Updated 3 years ago
- The code for paper Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models.☆13Apr 10, 2024Updated 2 years ago
- Ship a prebuilt Wine environment driven by box86 & box64 on Ubuntu Touch☆13Jul 12, 2026Updated last week
- A framework to meta-train transformers for causal ICL☆11Updated this week
- Exploring Model Kinship for Merging Large Language Models☆28Apr 16, 2025Updated last year
- Awesome-RL-Reasoning☆16May 31, 2026Updated last month
- [ACL 2025] Knowledge Unlearning for Large Language Models☆49Sep 18, 2025Updated 10 months ago
- This repository includes code and materials for the paper "Efficient PRM Training Data Synthesis via Formal Verification" (ACL 2026 Findi…☆18Apr 7, 2026Updated 3 months ago
- [ACL 2025] "CoT-UQ: Improving Response-wise Uncertainty Quantification in LLMs with Chain-of-Thought"☆17Apr 3, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official Repository of Paper "Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs"☆15Sep 25, 2025Updated 9 months ago
- Ruler: A Model-Agnostic Method to Control Generated Length for Large Language Models☆41Sep 30, 2024Updated last year
- ☆18Jun 2, 2026Updated last month
- Some papers on low-resource knowledge base population. (mostly from 2018 to 2020)☆20Mar 25, 2020Updated 6 years ago
- [NeurIPS 2024] TALoS: Enhancing Semantic Scene Completion via Test-time Adaptation on the Line of Sight☆37Feb 27, 2026Updated 4 months ago
- Make Your Training Flexible: Towards Deployment-Efficient Video Models☆40Jun 11, 2025Updated last year
- Port of GGML to C#☆13Jul 1, 2023Updated 3 years ago