Code for ICML 2024 paper
☆34Sep 18, 2025Updated 11 months ago
Alternatives and similar repositories for larimar
Users that are interested in larimar are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An implementation is provided here for the NeurIPS2024 paper "MemoryFormer : Minimize Transformer Computation by Removing Fully-Connected…☆16Mar 24, 2026Updated 5 months ago
- The official implementation of the paper "Self-Updatable Large Language Models by Integrating Context into Model Parameters"☆16May 18, 2025Updated last year
- Personalized Steering of Large Language Models: Versatile Steering Vectors Through Bi-directional Preference Optimization☆51Jul 28, 2024Updated 2 years ago
- Casande-RL☆11May 9, 2023Updated 3 years ago
- Code for Fast-weight Product Key Memory (FwPKM)☆22Mar 18, 2026Updated 5 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- analyse problems of AI with Math and Code☆31Jul 28, 2025Updated last year
- Applies ROME and MEMIT on Mamba-S4 models☆16Apr 5, 2024Updated 2 years ago
- ☆12Jun 5, 2024Updated 2 years ago
- Official code for the paper "Attention as a Hypernetwork"☆59Feb 24, 2026Updated 6 months ago
- Official repository of paper "RNNs Are Not Transformers (Yet): The Key Bottleneck on In-context Retrieval"☆27Apr 17, 2024Updated 2 years ago
- ☆15Sep 7, 2022Updated 3 years ago
- 在番茄时间中与意图保持连接☆15Updated this week
- ☆49Mar 15, 2025Updated last year
- Code for the EMNLP24 paper "A simple and effective L2 norm based method for KV Cache compression."☆19Dec 13, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Implementation of LaViC (KDD 2025)☆13Jun 1, 2025Updated last year
- [ICLR 2025] FLAT: LLM Unlearning via Loss Adjustment with Only Forget Data☆14Feb 26, 2025Updated last year
- [ACL 2025] Squeezed Attention: Accelerating Long Prompt LLM Inference☆58Nov 20, 2024Updated last year
- TheGlueNote is representation model for note-wise music alignment.☆14Jul 19, 2024Updated 2 years ago
- ☆26Jan 16, 2025Updated last year
- source code for EMNLP 2022 paper HEGEL: Hypergraph Transformer for Long Document Summarization☆15Oct 24, 2022Updated 3 years ago
- 机器学习部分算法实现,分类、聚类、回归(LR、Kmeans、GMM、PCA)☆10Mar 12, 2019Updated 7 years ago
- Code for the paper "Spectral Editing of Activations for Large Language Model Alignments"☆31Dec 20, 2024Updated last year
- Codebase for Context-aware Meta-learned Loss Scaling (CaMeLS). https://arxiv.org/abs/2305.15076.☆26Jan 23, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A small cli tool that downloads sheet music from MuseScore without the hassle☆15Oct 27, 2022Updated 3 years ago
- arXiv 2024 | ZIP: entropy-law data selection for efficient LLM alignment.☆28Jun 10, 2026Updated 2 months ago
- This is a reproduction of the paper 'Beyond Fully-Connected Layers with Quaternions: Parameterization of Hypercomplex Multiplications wit…☆13Aug 22, 2021Updated 5 years ago
- ☆17Oct 12, 2023Updated 2 years ago
- This is the accompanying repository to the paper - Automatic Estimation of Singing Voice Musical Dynamics☆16Oct 28, 2024Updated last year
- The official implementation of the paper **LVChat: Facilitating Long Video Comprehension**☆14Apr 15, 2024Updated 2 years ago
- Code for "Seeking Neural Nuggets: Knowledge Transfer in Large Language Models from a Parametric Perspective"☆33May 9, 2024Updated 2 years ago
- ☆18Aug 19, 2024Updated 2 years ago
- ☆18Aug 4, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Fork of github.com/UCSBarchlab/OpenTPU for the TGPTPU project☆15Jun 1, 2025Updated last year
- [NAACL 2025] Official Implementation of "HMT: Hierarchical Memory Transformer for Long Context Language Processing"☆80Mar 12, 2026Updated 5 months ago
- ☆15Jan 11, 2019Updated 7 years ago
- Official implementation of the paper "Pretraining Language Models to Ponder in Continuous Space"☆27Jul 21, 2025Updated last year
- [ICCV'23] Learning Vision-and-Language Navigation from YouTube Videos☆72Dec 27, 2024Updated last year
- Code and Data for "Language Modeling with Editable External Knowledge"☆39Jun 19, 2024Updated 2 years ago
- ☆41Mar 25, 2026Updated 5 months ago