β58Aug 10, 2026Updated last month
Alternatives and similar repositories for SWE-QA-Bench
Users that are interested in SWE-QA-Bench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SWE-Debate: Competitive Multi-Agent Debate for Software Issue Resolution [ICSE 2026]β34Nov 11, 2025Updated 10 months ago
- Must-read papers on Repository-level Code Generation & Issue Resolution π₯β330Aug 21, 2026Updated 3 weeks ago
- CodeRepoQA datasetβ15Feb 19, 2025Updated last year
- Code for our paper: "Building A Coding Assistant via Retrieval-Augmented Language Models"β10Nov 2, 2024Updated last year
- Advances and Frontiers of LLM-based Issue Resolution in Software Engineering A Comprehensive Surveyβ87Sep 2, 2026Updated 2 weeks ago
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- LongCodeZip: Compress Long Context for Code Language Models [ASE2025]β168May 12, 2026Updated 4 months ago
- β24Jun 2, 2026Updated 3 months ago
- Fork to run instances from SWE-rebenchβ30Jun 3, 2026Updated 3 months ago
- something for paper agentβ11Dec 18, 2024Updated last year
- Mixture of Expert (MoE) techniques for enhancing LLM performance through expert-driven prompt mapping and adapter combinations.β11Feb 11, 2024Updated 2 years ago
- β13Aug 9, 2023Updated 3 years ago
- Language Models for Code Completion: a Practical Evaluationβ13Jan 19, 2024Updated 2 years ago
- β10Apr 15, 2023Updated 3 years ago
- A framework for evaluating the effectiveness of chain-of-thought reasoning in language models.β19Feb 6, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding [ISSTA 2026]β32Feb 2, 2026Updated 7 months ago
- [AAAI 2025 oral] Evaluating Mathematical Reasoning Beyond Accuracyβ80Oct 9, 2025Updated 11 months ago
- Multi-Granularity LLM Debugger [ICSE2026]β101Jul 6, 2025Updated last year
- Multi-turn dataset management tool for LLM trainersβ13Mar 31, 2025Updated last year
- π UDS Software Factory Integration / Wayfinding Repoβ19Jan 7, 2026Updated 8 months ago
- "DeepResearch-Eval: An End-to-End Evaluation Framework for DeepResearch Systems"β51Oct 16, 2025Updated 11 months ago
- minimalistic AI library that resembles HF's transformersβ13Dec 31, 2024Updated last year
- [NeurIPS 2025] CodeCrash: Exposing LLM Fragility to Misleading Natural Language in Code Reasoningβ18Jan 24, 2026Updated 7 months ago
- [NeurIPS 2025 D&B] π SWE-bench Goes Live!β241Sep 10, 2026Updated last week
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- CoSearch: Joint Training of Reasoning and Document Ranking via Reinforcement Learning for Agentic Searchβ15Apr 28, 2026Updated 4 months ago
- Multi-source retrieval and function localization for repository repairβ37Jul 12, 2026Updated 2 months ago
- [CIKM'2024] "RecDiff: Diffusion Model for Social Recommendation"β90Jun 16, 2025Updated last year
- β13May 19, 2024Updated 2 years ago
- TeamTalkβ17Mar 25, 2015Updated 11 years ago
- [ASE 2025] CoSIL: Issue Localization via Iteritive Code Graph Searchingβ26May 31, 2026Updated 3 months ago
- [ACL25] FEA-Bench: A Benchmark for Evaluating Repository-Level Code Generation for Feature Implementationβ62Jan 28, 2026Updated 7 months ago
- "SALLM: Security Assessment of Generated Code" accepted at ASYDE workshop co-located with ASE'24.β17Jul 21, 2026Updated last month
- Semi-automated modelling and Model-Based Testing for CosmWasm contractsβ17Jun 28, 2024Updated 2 years ago
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Convert GitHub PRs into Harbor tasksβ83Jul 13, 2026Updated 2 months ago
- This is the tool released in ICSE 2024 paper "Domain Knowledge Matters: Improving Prompts with Fix Templates for Repairing Python Type Erβ¦β17Jun 5, 2023Updated 3 years ago
- "FastAgent: Simple, Fast, and Strong LLM Agents"β58Feb 10, 2026Updated 7 months ago
- ALAS: Autonomous Learning Agent Systemβ19Aug 14, 2025Updated last year
- [ASE2024] Mutual Learning-Based Framework for Enhancing Robustness of Code Models via Adversarial Trainingβ11Sep 13, 2024Updated 2 years ago
- β13Jun 27, 2025Updated last year
- [ICML'2024] "FlashST: A Simple and Universal Prompt-Tuning Framework for Traffic Prediction"β94Sep 2, 2024Updated 2 years ago