☆16Nov 20, 2025Updated 9 months ago
Alternatives and similar repositories for CoreCodeBench
Users that are interested in CoreCodeBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A iterative feedback driven benchmark on LLM's instruction following ability☆59May 25, 2026Updated 3 months ago
- [ICLR'26] R-HORIZON: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?☆26May 9, 2026Updated 3 months ago
- ☆14May 25, 2026Updated 3 months ago
- A minimal AI coding agent powered by Anthropic's Claude. Interactive terminal interface with tool execution. Visit https://lldong.github.…☆16Jul 15, 2025Updated last year
- wechat robot微信聊天机器人,扩展了查询天气预报,机器人聊天等功能☆13Jun 17, 2019Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆22Nov 5, 2024Updated last year
- ☆288May 13, 2026Updated 3 months ago
- [SIGIR 2025] Official impl. of "MRAMG-Bench: A Comprehensive Benchmark for Advancing Multimodal Retrieval-Augmented Multimodal Generation…☆19Apr 15, 2025Updated last year
- Enhancing contextual understanding in large language models through contrastive decoding☆19May 3, 2024Updated 2 years ago
- DAR introduces the diagonal scanning order for next-token prediction and proposes a direction-aware autoregressive transformer framework.☆19Apr 16, 2025Updated last year
- EMNLP 2024 Findings "Schema-Driven Information Extraction from Heterogeneous Tables"☆28Dec 5, 2024Updated last year
- ☆42Apr 7, 2026Updated 4 months ago
- [FSE'2026] PlayCoder: Making LLM-Generated GUI Code Playable☆46Apr 22, 2026Updated 4 months ago
- A pipeline using LLMs for Knowledge Engineering, combining knowledge probing and Wikidata entity mapping.☆37Dec 29, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆23Jun 2, 2026Updated 2 months ago
- ICLR 2026: Agent-X Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks☆44Apr 28, 2026Updated 4 months ago
- ☆25Aug 20, 2025Updated last year
- Source-level code analysis toolkit for SAST, context engineering, and AI coding☆34Jun 7, 2026Updated 2 months ago
- Hypothetical Minds is an autonomous LLM-based agent for diverse multi-agent settings, integrating a Theory of Mind module Theory of Mind …☆64Jul 13, 2024Updated 2 years ago
- [SIGIR 2023] Schema-aware Reference as Prompt Improves Data-Efficient Knowledge Graph Construction☆42Apr 5, 2023Updated 3 years ago
- [ICML 2024] Language Models Represent Beliefs of Self and Others☆37Sep 26, 2024Updated last year
- ☆32Dec 17, 2023Updated 2 years ago
- This is the official implementation of paper "Leveraging Dual Process Theory in Language Agent Framework for Simultaneous Human-AI Collab…☆61Nov 22, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆45Aug 20, 2025Updated last year
- Multi-SWE-bench: A Multilingual Benchmark for Issue Resolving☆359Dec 18, 2025Updated 8 months ago
- The official repository of the OpenToM dataset☆34Feb 2, 2025Updated last year
- [NeurIPS 2024] OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI☆106Mar 6, 2025Updated last year
- code for paper "Compositional Text-to-Image Synthesis with Attention Map Control of Diffusion Models"☆46Sep 21, 2023Updated 2 years ago
- An implementation of Laplacian surface editing.☆40Dec 21, 2019Updated 6 years ago
- Intro to using DSPy with Kuzu to enrich the data within the Nobel Laureate mentorship network☆16Sep 16, 2025Updated 11 months ago
- Multi-source retrieval and function localization for repository repair☆35Jul 12, 2026Updated last month
- [NeurIPS'25] Beyond Accuracy: Dissecting Mathematical Reasoning for LLMs Under Reinforcement Learning☆16Dec 12, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆18Jan 19, 2026Updated 7 months ago
- Code for the paper "Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns"☆18Mar 15, 2024Updated 2 years ago
- ☆16Aug 5, 2025Updated last year
- We release Open Meditron, a fully open, clinician-audited medical training corpus and evaluation protocol that closes the open-vs-closed …☆17Aug 3, 2026Updated 3 weeks ago
- [ICLR 2026] VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications☆167Feb 22, 2026Updated 6 months ago
- A benchmark to evaluate search-augmented LLMs☆17Aug 28, 2025Updated last year
- An MLX implementation of Meta AI's ESM-2 protein language model☆16Aug 16, 2025Updated last year