Fast Memorization of Prompt Improves Context Awareness of Large Language Models (Findings of EMNLP 2024)
☆22Oct 22, 2024Updated last year
Alternatives and similar repositories for FastMem
Users that are interested in FastMem are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for "In-Context Former: Lightning-fast Compressing Context for Large Language Model" (Findings of EMNLP 2024)☆21Nov 21, 2024Updated last year
- ☆14Oct 17, 2024Updated last year
- ☆17Jun 25, 2025Updated last year
- ☆23Dec 17, 2024Updated last year
- ☆41May 25, 2026Updated 3 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICLR 2025] Linear Combination of Saved Checkpoints Makes Consistency and Diffusion Models Better☆16Feb 15, 2025Updated last year
- Learnable Graph Discovery☆10May 17, 2019Updated 7 years ago
- ☆38Jan 17, 2025Updated last year
- ☆17Jun 12, 2018Updated 8 years ago
- Confidence-aware Personalized Federated Learning via Variational Expectation Maximization [Accepted at CVPR 2023]☆16Nov 8, 2023Updated 2 years ago
- Official Implementation of "DeCoRe: Decoding by Contrasting Retrieval Heads to Mitigate Hallucination"☆30Dec 18, 2024Updated last year
- [ACL 2025] Knowledge Unlearning for Large Language Models☆49Sep 18, 2025Updated last year
- Corresponding code to "Improving Robustness of ML Classifiers against Realizable Evasion Attacks Using Conserved Features" @ USENIX Secur…☆11Aug 5, 2019Updated 7 years ago
- A Model Agnostic function to directly remove specified layers from the LLM☆10May 23, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆11May 17, 2024Updated 2 years ago
- A toolkit for automated alignment research.☆16Jul 3, 2026Updated 2 months ago
- ☆79Mar 6, 2025Updated last year
- A Workbench for Autograding Retrieve/Generate Systems☆15Jun 30, 2025Updated last year
- ☆15Jul 24, 2022Updated 4 years ago
- ☆12Sep 22, 2024Updated last year
- ☆15Feb 8, 2025Updated last year
- Codes of BaiLian (POJ), Luogu, LeetCode & Course OJ☆16Dec 21, 2019Updated 6 years ago
- ☆48Apr 8, 2026Updated 5 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆57Nov 18, 2024Updated last year
- ☆15Jun 6, 2023Updated 3 years ago
- ☆23Jun 10, 2025Updated last year
- Code for the paper Xiangqi-R1: Enhancing Spatial Strategic Reasoning in LLMs for Chinese Chess via Reinforcement Learning☆15Jul 23, 2025Updated last year
- ☆64Jun 2, 2026Updated 3 months ago
- Library for minimizing Pseudo-Boolean functions☆27Oct 31, 2023Updated 2 years ago
- ☆13Oct 14, 2020Updated 5 years ago
- A very limited implementation of arXiv:1904.00759☆13Dec 2, 2019Updated 6 years ago
- [ICLR'26] SinkTrack: Attention Sink based Context Anchoring for Large Language Models☆19Apr 23, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- USTC 18 计算机学院 个人资源分享☆23Mar 7, 2021Updated 5 years ago
- Code for the paper "Evading Black-box Classifiers Without Breaking Eggs" [SaTML 2024]☆21Apr 15, 2024Updated 2 years ago
- [ACL'25] We propose a novel fine-tuning method, Separate Memory and Reasoning, which combines prompt tuning with LoRA.☆89Nov 2, 2025Updated 10 months ago
- Dataset and baseline for Coling 2022 long paper (oral): "ConFiguRe: Exploring Discourse-level Chinese Figures of Speech"☆12Jul 27, 2023Updated 3 years ago
- Source code of our paper "Focus on the Target’s Vocabulary: Masked Label Smoothing for Machine Translation" @ ACL 2022☆13Apr 13, 2022Updated 4 years ago
- Large Language Models in Molecular Embeddings☆12May 1, 2024Updated 2 years ago
- Vision Large Language Models trained on M3IT instruction tuning dataset☆17Aug 16, 2023Updated 3 years ago