Official repository for "NoLiMa: Long-Context Evaluation Beyond Literal Matching"
☆202Jul 17, 2025Updated last year
Alternatives and similar repositories for NoLiMa
Users that are interested in NoLiMa are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A framework for steering MoE models by detecting and controlling behavior-linked experts.☆36Sep 12, 2025Updated 10 months ago
- Official Implementation for the ICLR2023 paper "Fuzzy Alignments in Directed Acyclic Graph for Non-autoregressive Machine Translation"☆14Mar 1, 2023Updated 3 years ago
- FamilyBench evaluation tool for testing the relational reasoning capabilities of Large Language Models (LLMs).☆47May 4, 2026Updated 2 months ago
- ☆16Oct 28, 2025Updated 9 months ago
- Implementation for IceCache: Memory-Efficient KV-cache Management for Long-Sequence LLMs (ICLR 2026).☆20Jun 9, 2026Updated last month
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Locally hosted AI Agent Python Tool To Generate Novel Research Hypothesis + Titles + Abstracts☆30Apr 30, 2025Updated last year
- ☆58Oct 10, 2025Updated 9 months ago
- Code and Data for Evaluating the Evaluators☆16Aug 20, 2025Updated 11 months ago
- ☆19Oct 14, 2024Updated last year
- ☆27Jun 11, 2025Updated last year
- Easily take an entire YouTube playlist and turn it into high quality transcripts using Whisper.☆10Feb 1, 2024Updated 2 years ago
- This repo contains the source code for RULER: What’s the Real Context Size of Your Long-Context Language Models?☆1,589Jul 22, 2026Updated last week
- Running inference on the ZeroSCROLLS benchmark☆22Apr 18, 2024Updated 2 years ago
- [NeurIPS 2024] Fast Best-of-N Decoding via Speculative Rejection☆56Oct 29, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Homepage for ProLong (Princeton long-context language models) and paper "How to Train Long-Context Language Models (Effectively)"☆261Sep 12, 2025Updated 10 months ago
- ☆19Oct 18, 2025Updated 9 months ago
- An agentic CLI, built in rust.☆18Updated this week
- Lego for GRPO☆30May 27, 2025Updated last year
- Specialized AI agents for your bare metal. 100% On-Prem & Air-Gap ready.☆26Feb 21, 2026Updated 5 months ago
- [COLM'25] A Controlled Study on Long Context Extension and Generalization in LLMs☆65Mar 9, 2026Updated 4 months ago
- ☆20Aug 1, 2025Updated 11 months ago
- Sandboxed tools and JS runtime for AI agents☆17Jul 13, 2026Updated 2 weeks ago
- LongRoPE is a novel method that can extends the context window of pre-trained LLMs to an impressive 2048k tokens.☆290Oct 28, 2025Updated 9 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Lemonade Pi Agent CLI Plugin☆15Jun 9, 2026Updated last month
- Code for paper: Long cOntext aliGnment via efficient preference Optimization☆26Oct 10, 2025Updated 9 months ago
- A RAG system designed to process documents with multimodal content. It can generate factual, context-aware answers to user queries, based…☆26Dec 13, 2024Updated last year
- ☆54Jul 18, 2024Updated 2 years ago
- Positional Skip-wise Training for Efficient Context Window Extension of LLMs to Extremely Length (ICLR 2024)☆208May 20, 2024Updated 2 years ago
- A simple frontend page to interact with an OpenAI like API☆17Jan 31, 2025Updated last year
- JotItNow is a AI Voice Notes App☆26Mar 6, 2025Updated last year
- An experiment in using DuckDB for a datalog / egg☆27Oct 11, 2023Updated 2 years ago
- Cohere Toolkit is a collection of prebuilt components enabling users to quickly build and deploy RAG applications.☆30Jan 19, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- KV Cache Compression, But What Must We Give in Return? A Comprehensive Benchmark of Long Context Capable Approaches. EMNLP Findings 2024☆90Feb 27, 2025Updated last year
- iOS app. Work from home? Geostreaks motivates you to go outside every day.☆12Feb 24, 2025Updated last year
- ☆43Jun 14, 2026Updated last month
- A Python utility for building FalkorDB databases from CSV inputs☆15Jul 22, 2026Updated last week
- klmbr - a prompt pre-processing technique to break through the barrier of entropy while generating text with LLMs☆90Sep 22, 2024Updated last year
- A codebase for pretraining multi-billion-scale sparse GPTs.☆24Feb 9, 2026Updated 5 months ago
- [ICML‘2024] "LoCoCo: Dropping In Convolutions for Long Context Compression", Ruisi Cai, Yuandong Tian, Zhangyang Wang, Beidi Chen☆17Sep 7, 2024Updated last year