The official implementation of the paper "Self-Updatable Large Language Models by Integrating Context into Model Parameters"
☆15May 18, 2025Updated last year
Alternatives and similar repositories for SELF-PARAM
Users that are interested in SELF-PARAM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Complexity Based Prompting for Multi-Step Reasoning☆17Mar 10, 2023Updated 3 years ago
- ☆12Jun 5, 2024Updated 2 years ago
- [KDD 2023] code for "Test accuracy vs. generalization gap: model selection in NLP without accessing training or testing data" https://arx…☆12Oct 17, 2022Updated 3 years ago
- Code for "Can Retriever-Augmented Language Models Reason? The Blame Game Between the Retriever and the Language Model", EMNLP Findings 20…☆28Nov 2, 2023Updated 2 years ago
- TBC☆28Nov 2, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The official implementation of the paper "Mem-α: Learning Memory Construction via Reinforcement Learning"☆218Dec 25, 2025Updated 7 months ago
- FocusLLM: Scaling LLM’s Context by Parallel Decoding☆45Dec 8, 2024Updated last year
- [arxiv: 2503.23895] Dynamic Parametric Retrieval Augmented Generation for Test-time Knowledge Enhancement☆182Aug 14, 2025Updated 11 months ago
- Code for ICML 2024 paper☆34Sep 18, 2025Updated 10 months ago
- Code for "Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous…☆46May 18, 2026Updated 2 months ago
- MinPrompt: Graph-based Minimal Prompt Data Augmentation for Few-shot Question Answering☆14May 3, 2024Updated 2 years ago
- ☆41Nov 30, 2023Updated 2 years ago
- Official source code for AAAI 2025 paper: CoRA: Collaborative Information Perception by Large Language Model's Weights for Recommendatio…☆18Dec 11, 2024Updated last year
- A curated reading list on harness engineering for recursive self-improvement of LLM agents (EN/ZH).☆19Jul 9, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆24Feb 18, 2025Updated last year
- ☆18Jan 17, 2024Updated 2 years ago
- ☆12Oct 17, 2022Updated 3 years ago
- ☆26Mar 4, 2025Updated last year
- ☆17Jun 21, 2024Updated 2 years ago
- [NLPCC 2022] Kformer: Knowledge Injection in Transformer Feed-Forward Layers☆39Oct 20, 2022Updated 3 years ago
- [EMNLP 2024 (Oral)] Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA☆155Dec 22, 2025Updated 7 months ago
- ☆24May 19, 2023Updated 3 years ago
- Official Repository for "Hyper-CL: Conditioning Sentence Representations with Hypernetworks"☆16Jun 3, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- The code of AAAI 2020 paper "Transparent Classification with Multilayer Logical Perceptrons and Random Binarization".☆23Mar 10, 2024Updated 2 years ago
- Materials for "Multi-property Steering of Large Language Models with Dynamic Activation Composition"☆14Nov 22, 2024Updated last year
- ☆18Mar 25, 2021Updated 5 years ago
- ☆22Feb 14, 2023Updated 3 years ago
- An easy tool to extract slides from presentations ( lectures 😉 )☆14Dec 10, 2023Updated 2 years ago
- ☆10Oct 25, 2024Updated last year
- ☆22Jun 16, 2025Updated last year
- A method for evaluating the high-level coherence of machine-generated texts. Identifies high-level coherence issues in transformer-based …☆12Mar 18, 2023Updated 3 years ago
- Open-Theatre: An Open-Source Toolkit for LLM-based Interactive Drama☆28Oct 20, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- (ACL 2025) Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation☆12May 21, 2025Updated last year
- fast trainer for educational purposes☆26Updated this week
- ☆25Dec 12, 2025Updated 7 months ago
- A plug-in of Microsoft DeepSpeed to fix the bug of DeepSpeed pipeline☆25Apr 16, 2021Updated 5 years ago
- [NeurIPS 2021] code for "Taxonomizing local versus global structure in neural network loss landscapes" https://arxiv.org/abs/2107.11228☆20Jan 7, 2022Updated 4 years ago
- Collections of Undergraduate Course Projects☆22Jul 17, 2026Updated last week
- Implementation of NAACL 2024 Outstanding Paper "LM-Infinite: Simple On-the-Fly Length Generalization for Large Language Models"☆153Mar 13, 2025Updated last year