DependEval: a hierarchical benchmark for evaluating LLMs on repository-level code understanding across 8 programming languages.
☆16Jul 28, 2025Updated last year
Alternatives and similar repositories for DependEval
Users that are interested in DependEval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Feb 2, 2026Updated 6 months ago
- This is a JLU-SNL-COMPILER project☆13Oct 6, 2023Updated 2 years ago
- ☆23Mar 28, 2026Updated 4 months ago
- Dataset of Codex generated tests for the CodaMosa project☆19Jun 2, 2023Updated 3 years ago
- The open-source repository for PAL: Sample-Efficient Personalized Reward Modeling for Pluralistic Alignment, which provides a general per…☆17Aug 28, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code to reproduce results of our experiments using LoRe☆17Jun 10, 2026Updated 2 months ago
- ☆22Dec 28, 2024Updated last year
- CodeMind is a generic framework for evaluating inductive code reasoning of LLMs. It is equipped with a static analysis component that ena…☆42Feb 18, 2026Updated 5 months ago
- ☆28Nov 13, 2024Updated last year
- [KDD 2025] Fine-tuning Multimodal Large Language Models for Product Bundling☆16Sep 20, 2025Updated 10 months ago
- CONCOCTION is an automated machine learning-based vulnerability detection framework that combines static source code information and dyna…☆28Aug 18, 2024Updated last year
- This is an evaluation set for the problem of directed/targeted test input generation. We use it to benchmark the ability of Large Languag…☆34Mar 11, 2025Updated last year
- ☆36Jan 27, 2025Updated last year
- Python library for backtranslation (with Google Translate)☆12Jan 11, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- SWE-Debate: Competitive Multi-Agent Debate for Software Issue Resolution [ICSE 2026]☆33Nov 11, 2025Updated 9 months ago
- ☆15Oct 17, 2023Updated 2 years ago
- The implementation of paper "Strategy-aware Bundle Recommender System", SIGIR'23.☆17Sep 4, 2023Updated 2 years ago
- Code and data for the paper "Understanding Hidden Context in Preference Learning: Consequences for RLHF"☆27Aug 21, 2024Updated last year
- Offical implementation of our paper "Exploring the Potential of Diffusion Large Language Models in Code Generation".☆23Oct 29, 2025Updated 9 months ago
- This repository contains the code for applying One-Token Approximation to a pretrained language model using subword-level tokenization.☆12May 7, 2020Updated 6 years ago
- This repository contains the WordNet Language Model Probing (WNLaMPro) dataset introduced in "Rare Words: A Major Problem for Contextuali…☆14Feb 2, 2020Updated 6 years ago
- ☆13Oct 17, 2020Updated 5 years ago
- ☆13Jul 6, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The implementation of paper "EliMRec: Eliminating single-modal bias in multimedia recommendation", MM'22.☆24Dec 7, 2023Updated 2 years ago
- [NAACL 2025] Benchmark for Repository-Level Code Generation, focus on Executability, Correctness from Test Cases and Usage of Contexts fr…☆46Jan 8, 2026Updated 7 months ago
- SocialDial: A Benchmark for Socially-Aware Dialogue Systems (SIGIR'23)☆16Aug 4, 2023Updated 3 years ago
- a small, non-commercial, fair-use subset of the Penn-Treebank, in JSON.☆17Apr 10, 2018Updated 8 years ago
- Code and data for "KoDialogBench: Evaluating Conversational Understanding of Language Models with Korean Dialogue Benchmark" (LREC-COLING…☆18Apr 15, 2025Updated last year
- ☆29Jun 2, 2026Updated 2 months ago
- [EMNLP 2021] Code and data for our paper "Vision-and-Language or Vision-for-Language? On Cross-Modal Influence in Multimodal Transformers…☆20Jan 17, 2022Updated 4 years ago
- Data and preprocessing scripts for SemEval 2022 Task 2: Multilingual Idiomaticity Detection and Sentence Embedding☆16Feb 3, 2022Updated 4 years ago
- ☆18Jan 17, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆14May 7, 2019Updated 7 years ago
- 汇总了关于在 吉林大学 生存所需的相关仓库☆428Jun 24, 2026Updated last month
- ☆31Jul 4, 2026Updated last month
- MutAP: A prompt_based learning technique to automatically generate test cases with Large Language Model☆57Mar 7, 2025Updated last year
- RESTful web service for open-korean-text☆19Oct 17, 2021Updated 4 years ago
- ☆20Oct 22, 2021Updated 4 years ago
- ProxySR (Unsupervised Proxy Selection for Session-based Recommender Systems, SIGIR'21)☆13Nov 16, 2021Updated 4 years ago