☆54Oct 24, 2024Updated last year
Alternatives and similar repositories for nocha
Users that are interested in nocha are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12Sep 1, 2021Updated 4 years ago
- ☆61Sep 24, 2024Updated last year
- ☆29Dec 2, 2024Updated last year
- Suri: Multi-constraint instruction following for long-form text generation [EMNLP’24]☆27Oct 3, 2025Updated 10 months ago
- ☆22Sep 19, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆12Feb 21, 2021Updated 5 years ago
- Repository for DEMETR: Diagnosing Evaluation Metrics for Translation☆17Nov 29, 2022Updated 3 years ago
- Homepage for ProLong (Princeton long-context language models) and paper "How to Train Long-Context Language Models (Effectively)"☆263Sep 12, 2025Updated 11 months ago
- Yet another frontend for LLM, written using .NET and WinUI 3☆11Sep 14, 2025Updated 11 months ago
- ☆26Dec 12, 2025Updated 8 months ago
- The HELMET Benchmark☆225Apr 17, 2026Updated 3 months ago
- GPT* - Training faster small transformers using ALiBi, Parallel Residual Connections and more!☆20Oct 29, 2022Updated 3 years ago
- Exploring limitations of LLM-as-a-judge☆20Aug 17, 2024Updated last year
- ☆27Jun 4, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICLR 2025] Monet: Mixture of Monosemantic Experts for Transformers☆79Jun 23, 2025Updated last year
- DiffusER: Discrete Diffusion via Edit-based Reconstruction (Reid, Hellendoorn & Neubig, 2022)☆55Aug 3, 2025Updated last year
- an auto-sleeping and -waking framework around llama.cpp☆13Feb 8, 2025Updated last year
- Official implementation for "Law of the Weakest Link: Cross capabilities of Large Language Models"☆43Oct 1, 2024Updated last year
- Learning from Mixed Rollouts: Logit Fusion as a Bridge Between Imitation and Exploration☆17Feb 24, 2026Updated 5 months ago
- ☆15Jul 1, 2020Updated 6 years ago
- A controlled benchmark on evaluating and studying the dynamics of Long Context Language Models☆26Oct 17, 2025Updated 9 months ago
- REBUS: A Robust Evaluation Benchmark of Understanding Symbols☆13Aug 13, 2024Updated 2 years ago
- AI_Powered_Dev_Search_Engine☆12Mar 10, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆13Jan 30, 2023Updated 3 years ago
- ☆58Jun 2, 2026Updated 2 months ago
- [EMNLP 2023] Question Answering as Programming for Solving Time-Sensitive Questions☆12Dec 18, 2023Updated 2 years ago
- Evaluating LLMs with fewer examples☆182Jul 4, 2026Updated last month
- Repository for ACL'22 paper: Dynamic Latent Extraction for Abstractive Long-Input Summarization☆57Aug 2, 2023Updated 3 years ago
- This repo contains the source code for RULER: What’s the Real Context Size of Your Long-Context Language Models?☆1,604Jul 22, 2026Updated 3 weeks ago
- ☆17Apr 7, 2025Updated last year
- Continual Memorization of Factoids in Large Language Models☆12Nov 20, 2024Updated last year
- ☆57Aug 10, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official repository for "PostMark: A Robust Blackbox Watermark for Large Language Models"☆29Aug 30, 2024Updated last year
- BABILong is a benchmark for LLM evaluation using the needle-in-a-haystack approach.☆253Jun 1, 2026Updated 2 months ago
- Code & data for EMNLP 2020 paper "MOCHA: A Dataset for Training and Evaluating Reading Comprehension Metrics".☆16May 3, 2022Updated 4 years ago
- [EMNLP2025] Remedy: Learning Machine Translation Evaluation from Human Preferences with Reward Modeling☆16Nov 20, 2025Updated 8 months ago
- An original implementation of the paper "CREPE: Open-Domain Question Answering with False Presuppositions"☆16Nov 5, 2024Updated last year
- ☆18May 18, 2025Updated last year
- The first Object-Oriented Programming (OOP) Evaluation Benchmark for LLMs☆27Jan 15, 2025Updated last year