LatentMAS with kNN kv cache pruning | up to 40% more memory efficient and 30% faster
☆19Dec 10, 2025Updated 8 months ago
Alternatives and similar repositories for kNN-latentMAS
Users that are interested in kNN-latentMAS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Code for Lever LM: Configuring In-Context Sequence to Lever Large Vision Language Models☆18Oct 4, 2024Updated last year
- [ICML 2026 Spotlight] Latent Collaboration in Multi-Agent Systems☆1,079Jun 18, 2026Updated last month
- CopT: Contrastive On-Policy Thinking with Continuous Spaces for General and Agentic Reasoning☆18May 21, 2026Updated 2 months ago
- ☆77May 8, 2026Updated 3 months ago
- CoFiRec: Coarse-to-Fine Tokenization for Generative Recommendationn☆21Jan 23, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- "Parallel Test-Time Scaling for Latent Reasoning Models"☆24Apr 12, 2026Updated 3 months ago
- [NeurIPS 2024 D&B] Evaluating Copyright Takedown Methods for Language Models☆17Jul 17, 2024Updated 2 years ago
- C++ port of the Decay0/GENBB fortran Monte Carlo code for the generation of standard decay or double beta decay processes for various rad…☆15Dec 18, 2025Updated 7 months ago
- ☆11Dec 8, 2022Updated 3 years ago
- ☆19Oct 6, 2025Updated 10 months ago
- ☆14Oct 11, 2023Updated 2 years ago
- ☆13Apr 12, 2026Updated 3 months ago
- 3D-LMVIC: Learning-based Multi-View Image Coding with 3D Gaussian Geometric Priors☆15Jun 19, 2025Updated last year
- ☆15Jun 6, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆15Feb 18, 2025Updated last year
- [CCS 2026] The official implementation of our CCS 2026 paper "ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathological…☆16Updated this week
- ☆11Oct 26, 2022Updated 3 years ago
- ☆11Jun 27, 2022Updated 4 years ago
- Automatic Jailbreaking of the Text-to-Image Generative AI Systems☆15Jun 23, 2024Updated 2 years ago
- [USENIX Security 2022] Mitigating Membership Inference Attacks by Self-Distillation Through a Novel Ensemble Architecture☆16Aug 29, 2022Updated 3 years ago
- [ICLR'26] The official code implementation for "Cache-to-Cache: Direct Semantic Communication Between Large Language Models"☆426Mar 13, 2026Updated 4 months ago
- Ensemble Neural Representation Networks☆12Jan 5, 2022Updated 4 years ago
- Code for our NeurIPS 2023 paper Towards Evaluating Transfer-based Attacks Systematically, Practically, and Fairly☆14Jan 22, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ACL 2026] RAG over Tables: Hierarchical Memory Index, Multi-Stage Retrieval, and Benchmarking.☆32Oct 7, 2025Updated 10 months ago
- SEU Summer School project, based on Kotlin and Java.☆12Sep 15, 2023Updated 2 years ago
- ☆12Aug 3, 2020Updated 6 years ago
- ☆33Apr 1, 2026Updated 4 months ago
- Multi-task learning with compressible features for collaborative intelligence☆10May 15, 2019Updated 7 years ago
- 极简、纯粹、黑白风格的 Hexo 博客主题。☆16Jan 14, 2026Updated 6 months ago
- Official code of paper:SADN: Learned Light Field Image Compression with Spatial-Angular Decorrelation☆15May 6, 2023Updated 3 years ago
- ☆11Jun 24, 2021Updated 5 years ago
- Codebase for SIGNET: Efficient Neural Representations for Light Fields☆15Jul 27, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Learning ReLU INRs with B-spline wavelets.☆14Jun 5, 2024Updated 2 years ago
- Go implementation of Recursive Language Models (RLM) - inference-time scaling for arbitrarily long contexts☆19May 12, 2026Updated 2 months ago
- 🐙「久美子大冒险」2D像素小游戏☆21Oct 10, 2024Updated last year
- ☆11May 28, 2023Updated 3 years ago
- ☆22Dec 14, 2023Updated 2 years ago
- Official PyTorch implementation for "Towards Lightweight Controllable Audio Synthesis with Conditional Implicit Neural Representations".☆21Dec 3, 2021Updated 4 years ago
- Meet IFR: a bio-inspired engine solving RAG’s biggest flaws. It achieves true O(1) scaling latency stays <5ms even as data grows 1000x. W…☆15Apr 3, 2026Updated 4 months ago