Homepage for paper “MeKi : Memory-based Expert Knowledge Injection for Efficient LLM Scaling”
☆29Mar 5, 2026Updated 5 months ago
Alternatives and similar repositories for MeKi
Users that are interested in MeKi are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An implementation is provided here for the NeurIPS2024 paper "MemoryFormer : Minimize Transformer Computation by Removing Fully-Connected…☆16Mar 24, 2026Updated 4 months ago
- ☆15Sep 25, 2025Updated 10 months ago
- MiSS is a novel PEFT method that features a low-rank structure but introduces a new update mechanism distinct from LoRA, achieving an exc…☆35Mar 9, 2026Updated 5 months ago
- [NeurIPS 2024] The official implementation of "Kangaroo: Lossless Self-Speculative Decoding for Accelerating LLMs via Double Early Exitin…☆73Jun 26, 2024Updated 2 years ago
- Pure C wrapper library to use llama.cpp with Linux and Windows as simple as possible.☆15Jul 28, 2026Updated last week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Experimental interface environment for open source LLM, designed to democratize the use of AI. Powered by llama-cpp, llama-cpp-python and…☆19Oct 11, 2025Updated 9 months ago
- My Implementation of Q-Sparse: All Large Language Models can be Fully Sparsely-Activated☆37Aug 14, 2024Updated last year
- ☆13Feb 17, 2025Updated last year
- Spartan is an algorithm for training sparse neural network models. This repository accompanies the paper "Spartan Differentiable Sparsity…☆26Oct 31, 2022Updated 3 years ago
- This is a fork of SGLang for hip-attention integration. Please refer to hip-attention for detail.☆18Mar 31, 2026Updated 4 months ago
- Mic-controlled mouse clicks☆17Oct 6, 2025Updated 10 months ago
- DataBaseLab,XJTU 西交数据库实验☆14Jun 25, 2024Updated 2 years ago
- Milk-V Duo. Access to Internet throw USB RNDIS connection to host machine☆16Jan 11, 2024Updated 2 years ago
- Code for EACL 26 Findings paper "I-MCTS: Enhancing Agentic AutoML via Introspective Monte Carlo Tree Search"☆13Jan 28, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A lightweight kmer-based algorithm for designing diagnostic CRISPR assays using genome data.☆13Aug 19, 2024Updated last year
- ☆30Nov 29, 2025Updated 8 months ago
- GUI tool to QLoRA/LoRA-fine-tune LLMs and deploy to Ollama. Broad GPU support (NVIDIA/AMD/Intel/Apple) + CPU fallback.☆14Feb 18, 2026Updated 5 months ago
- mechanical stage☆12Aug 6, 2023Updated 3 years ago
- (Not actively updating)Vision Transformer Accelerator implemented in Vivado HLS for Xilinx FPGAs.☆26Dec 29, 2024Updated last year
- [ECCV 2024] Official Implementation of CoPT: Unsupervised Domain Adaptive Segmentation using Domain-Agnostic Text Embeddings☆10Feb 24, 2025Updated last year
- tensorrt部署教程☆11Aug 1, 2025Updated last year
- ☆45Jan 30, 2026Updated 6 months ago
- ☆77May 12, 2026Updated 2 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ICLR 2022] Training L_inf-dist-net with faster acceleration and better training strategies☆22Mar 16, 2022Updated 4 years ago
- [NeurIPS'25] VidEmo: Affective-Tree Reasoning for Emotion-Centric Video Foundation Models☆15Dec 7, 2025Updated 8 months ago
- Self-hosted voice for coding agents. Talk from any browser or a Telegram call, interrupt mid-sentence, clone any voice, and hand real wor…☆27Aug 3, 2026Updated last week
- Metrics for evaluating biological sequence design☆15Jul 22, 2026Updated 2 weeks ago
- pyPFC: An Open-Source Python Package for Phase Field Crystal Simulations☆16Jun 25, 2026Updated last month
- ☆33Jun 22, 2024Updated 2 years ago
- ☆16Dec 10, 2025Updated 8 months ago
- The original Shared Recurrent Memory Transformer implementation☆36Jul 11, 2025Updated last year
- ☆15Aug 26, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Mem-T: Densifying Rewards for Long-Horizon Memory Agents☆39Mar 22, 2026Updated 4 months ago
- Open source codebase for PRBench☆18Jan 15, 2026Updated 6 months ago
- ☆19Jan 29, 2026Updated 6 months ago
- The code for "AttentionPredictor: Temporal Pattern Matters for Efficient LLM Inference", Qingyue Yang, Jie Wang, Xing Li, Zhihai Wang, Ch…☆29Jul 15, 2025Updated last year
- ☆91Jan 10, 2026Updated 7 months ago
- Bloat Free, Portable and Lightweight LLM Frontend (Single HTML file). With Lorebook, Web Search, Macro Engine etc.☆22Aug 1, 2026Updated last week
- Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents☆31Apr 16, 2026Updated 3 months ago