☆46Jun 2, 2026Updated 3 months ago
Alternatives and similar repositories for LoCoBench
Users that are interested in LoCoBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LoCoBench-Agent: An Interactive Benchmark for LLM Agents in Long-Context Software Engineering☆22Jun 2, 2026Updated 3 months ago
- [CHIL 2024] Interpretation of Intracardiac Electrograms Through Textual Representations☆12Sep 4, 2024Updated 2 years ago
- Evaluating Durability: Benchmark Insights into Multimodal Watermarking☆12Jun 7, 2024Updated 2 years ago
- ☆31Apr 7, 2026Updated 5 months ago
- ☆69Jun 2, 2026Updated 3 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- This is the official PyTorch implementation for the paper "Gated Associative Memory: A Parallel O(N) Architecture for Efficient Sequence …☆16Sep 3, 2025Updated last year
- Official implementation of the ΔBelief-RL method.☆31Feb 28, 2026Updated 6 months ago
- The official implementation of EMNLP 2021 paper "#HowYouTagTweets: Learning User Hashtagging Preferences via Personalized Topic Attention…☆11Feb 21, 2023Updated 3 years ago
- Graphs and grammars for Context-Free Path Querying algorithms evaluation.☆11Updated this week
- Knowledge Graph based Question Answering benchmark.☆10Feb 1, 2020Updated 6 years ago
- ☆41Aug 2, 2026Updated last month
- Functional Optimal Transport: Map Estimation and Domain Adaptation for Functional data☆28Jun 7, 2021Updated 5 years ago
- [TrustNLP@NAACL 2025] BiasEdit: Debiasing Stereotyped Language Models via Model Editing☆18Sep 30, 2025Updated 11 months ago
- [ACL 2023]: Training Trajectories of Language Models Across Scales https://arxiv.org/pdf/2212.09803.pdf☆25Nov 14, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Dataset and model in the paper "SciXGen: A Scientific Paper Dataset for Context-Aware Text Generation"☆13Feb 14, 2022Updated 4 years ago
- The official repository for Trust-Region Adaptive Policy Optimization (TRAPO) – a novel hybrid framework designed to enhance large langua…☆16Mar 2, 2026Updated 6 months ago
- ☆20Apr 9, 2025Updated last year
- Nexusflow function call, tool use, and agent benchmarks.☆28Dec 13, 2024Updated last year
- Code Intelligence Engine — indexes your codebase and gives AI assistants deep understanding via MCP (semantic search, call graphs, 20+ to…☆19Feb 14, 2026Updated 6 months ago
- ☆15Nov 18, 2025Updated 9 months ago
- [ISSTA'25] A GitHub issue resolution benchmark with multi-aspect diversity in programming languages, repository domains and modality of i…☆17Jun 13, 2025Updated last year
- ☆23May 14, 2026Updated 3 months ago
- Code for Deep learning models for electrocardiograms are susceptible to adversarial attack☆23Feb 4, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Research on Complaints in Social Media (ACL 2019)☆15Aug 15, 2019Updated 7 years ago
- ☆29Jul 20, 2024Updated 2 years ago
- Symbol-Equivariant Recurrent Reasoning Model☆17Mar 4, 2026Updated 6 months ago
- [EMNLP 2023] An Empirical Exploration of Cross-domain Alignment between Language and Electroencephalogram☆31Nov 9, 2023Updated 2 years ago
- Measuring the Impact of Early-2025 AI on Experienced Open-Source Developer Productivity: https://metr.org/blog/2025-07-10-early-2025-ai-e…☆17Feb 23, 2026Updated 6 months ago
- ProgQuery is a system to extract useful syntactic and semantic information from source code programs and store it in a graph database for…☆17Jan 22, 2025Updated last year
- Generating SpartQA dataset☆16May 3, 2023Updated 3 years ago
- The collection of Context-Free Path Querying algorithms☆14Dec 16, 2025Updated 8 months ago
- TAT-DQA: Towards Complex Document Understanding By Discrete Reasoning☆26Sep 17, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Reproducing R1 for Code with Reliable Rewards☆13Apr 9, 2025Updated last year
- LoongRL: Reinforcement Learning for Advanced Reasoning over Long Contexts (ICLR 2026 Oral)☆36Feb 20, 2026Updated 6 months ago
- ☆21Jul 28, 2022Updated 4 years ago
- RACE is a multi-dimensional benchmark for code generation that focuses on Readability, mAintainability, Correctness, and Efficiency.☆14Oct 12, 2024Updated last year
- ☆32Sep 23, 2025Updated 11 months ago
- SSRL: Self-Search Reinforcement Learning☆212Aug 20, 2025Updated last year
- [NeurIPS 2023 D&B Track] Code and data for paper "Revisiting Out-of-distribution Robustness in NLP: Benchmarks, Analysis, and LLMs Evalua…☆37Jun 8, 2023Updated 3 years ago