☆18Dec 2, 2024Updated last year
Alternatives and similar repositories for UltraGist
Users that are interested in UltraGist are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Oct 3, 2024Updated last year
- This is the official repo of "QuickLLaMA: Query-aware Inference Acceleration for Large Language Models"☆54Jul 16, 2024Updated 2 years ago
- ☆12Apr 29, 2024Updated 2 years ago
- ☆37Oct 4, 2025Updated 10 months ago
- Paper List for Agentic Coding☆19May 22, 2026Updated 2 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- The repo for In-context Autoencoder☆178May 11, 2024Updated 2 years ago
- Code for the EMNLP24 paper "A simple and effective L2 norm based method for KV Cache compression."☆19Dec 13, 2024Updated last year
- Know2BIO: A Comprehensive Dual-View Benchmark for Evolving Biomedical Knowledge Graphs☆15Feb 10, 2026Updated 6 months ago
- KV cache compression via sparse coding☆18Oct 26, 2025Updated 9 months ago
- Official code repo for paper "Great Memory, Shallow Reasoning: Limits of kNN-LMs"☆24Apr 30, 2025Updated last year
- IKEA: Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent☆71May 13, 2025Updated last year
- The official repository of paper "Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory Framework".☆18Sep 22, 2025Updated 10 months ago
- Official code and resources for the paper "EXIT: Context-Aware Extractive Compression for Enhancing Retrieval-Augmented Generation."☆26Jul 15, 2026Updated last month
- Submodule of evalverse forked from [google-research/instruction_following_eval](https://github.com/google-research/google-research/tree/m…☆15May 4, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆13Oct 19, 2023Updated 2 years ago
- Efficient retrieval head analysis with triton flash attention that supports topK probability☆13Jun 15, 2024Updated 2 years ago
- [COLM'25] A Controlled Study on Long Context Extension and Generalization in LLMs☆65Mar 9, 2026Updated 5 months ago
- official github repository of Subdora python package☆12Nov 13, 2024Updated last year
- [EMNLP'24] LongHeads: Multi-Head Attention is Secretly a Long Context Processor☆32Apr 8, 2024Updated 2 years ago
- LongProc: Benchmarking Long-Context Language Models on Long Procedural Generation☆36Feb 26, 2026Updated 5 months ago
- ☆11Aug 4, 2024Updated 2 years ago
- [ICCV2025] V2PE: Improving Multimodal Long-Context Capability of Vision-Language Models with Variable Visual Position Encoding☆60Apr 4, 2026Updated 4 months ago
- mstar: Optimizing memory architecture for every LLM task as executable Python code.☆16Jun 12, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Rationale-enhanced language models are better continual relation learners (EMNLP 2023 Main Conference)☆12Oct 11, 2023Updated 2 years ago
- [ICML'24] Data and code for our paper "Training-Free Long-Context Scaling of Large Language Models"☆450Oct 16, 2024Updated last year
- FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions☆57Jul 3, 2024Updated 2 years ago
- Homepage for ProLong (Princeton long-context language models) and paper "How to Train Long-Context Language Models (Effectively)"☆263Sep 12, 2025Updated 11 months ago
- An efficient implementation of the NSA (Native Sparse Attention) kernel☆134Jun 24, 2025Updated last year
- [ACL 2024] Long-Context Language Modeling with Parallel Encodings☆169Jun 13, 2024Updated 2 years ago
- ACL 2024 | LooGLE: Long Context Evaluation for Long-Context Language Models☆200Aug 6, 2026Updated last week
- ☆11Jun 7, 2023Updated 3 years ago
- Code and data for Distributional Correlation–Aware Knowledge Distillation for Stock Trading Volume Prediction (ECML-PKDD 22)☆16Sep 6, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Tcurtsni: Reverse Instruction Chat, ever wonder what your LLM wants to ask you?☆23Jun 25, 2024Updated 2 years ago
- ☆19Jun 29, 2025Updated last year
- ☆29Apr 7, 2024Updated 2 years ago
- ☆27May 12, 2026Updated 3 months ago
- 知予人工智能:从学习者到研究者☆14Jan 20, 2025Updated last year
- FocusLLM: Scaling LLM’s Context by Parallel Decoding☆45Dec 8, 2024Updated last year
- 国家统计局中国省市县乡村5级地址抓取,http://www.stats.gov.cn/tjsj/tjbz/tjyqhdmhcxhfdm/2018/index.html☆12Jan 8, 2020Updated 6 years ago