☆18Mar 28, 2026Updated 4 months ago
Alternatives and similar repositories for SWE-Compass
Users that are interested in SWE-Compass are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆86Jun 19, 2026Updated last month
- ☆16Nov 1, 2025Updated 9 months ago
- SWE-Swiss: A Multi-Task Fine-Tuning and RL Recipe for High-Performance Issue Resolution☆105Sep 24, 2025Updated 10 months ago
- An implementation of online data mixing for the Pile dataset, based on the GPT-NeoX library.☆14Jan 9, 2024Updated 2 years ago
- [ACL 2024 Findings] The official repo for "ConceptMath: A Bilingual Concept-wise Benchmark for Measuring Mathematical Reasoning of Large …☆26May 29, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Official repository for our paper "FullStack Bench: Evaluating LLMs as Full Stack Coders"☆121May 7, 2025Updated last year
- ☆42Jul 15, 2025Updated last year
- Official Code of MEnvAgent☆25Feb 3, 2026Updated 6 months ago
- [FSE'2026] SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarks☆184May 12, 2026Updated 2 months ago
- [ICML 2026] SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark☆22May 6, 2026Updated 3 months ago
- ☆39Updated this week
- ☆13Mar 5, 2025Updated last year
- ☆53Oct 28, 2025Updated 9 months ago
- The Source Code for WebCompass☆22May 2, 2026Updated 3 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [ACL25' Findings] SWE-Dev is an SWE agent with a scalable test case construction pipeline.☆65Jul 21, 2025Updated last year
- ☆51Aug 5, 2025Updated last year
- ☆23May 7, 2026Updated 3 months ago
- ☆153May 13, 2026Updated 2 months ago
- C^3-Bench: The Things Real Disturbing LLM based Agent in Multi-Tasking☆38Mar 1, 2026Updated 5 months ago
- ☆10Nov 14, 2024Updated last year
- a new cfi mechanism☆33Sep 23, 2021Updated 4 years ago
- Simple code for the tutorial on Polynomial Nets.☆13Jan 19, 2023Updated 3 years ago
- SWE-Flow: Synthesizing Software Engineering Data in a Test-Driven Manner☆40Jun 29, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆48Dec 12, 2024Updated last year
- An AI project to provide `private` chat and RAG service. 一个提供私有化检索增强生成的AI项目☆11Jul 14, 2024Updated 2 years ago
- Multi-agent synthetic data generation pipeline capable of generating and validating long horizon terminal/coding tasks for RL training☆72Jul 28, 2025Updated last year
- SimpleVQA: Multimodal Factuality Evaluation for Multimodal Large Language Models☆15Feb 20, 2025Updated last year
- ☆44Jul 23, 2026Updated 2 weeks ago
- [ICML 2026 Oral] Agent-native Mid-training for Software Engineering☆73Jun 7, 2026Updated 2 months ago
- ✨ RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems - ICLR 2024☆214Aug 16, 2024Updated last year
- Qwen1.5大模型微调、基于PEFT框架LoRA微调,在数据集HC3-Chinese上实现文本分类。☆12Jun 29, 2024Updated 2 years ago
- Heuristic filtering framework for RefineCode☆88Mar 13, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ACL 2025] Graph Aligned Large Language Models for Improved Source Code Understanding☆44May 18, 2025Updated last year
- [ISSTA'25] A GitHub issue resolution benchmark with multi-aspect diversity in programming languages, repository domains and modality of i…☆17Jun 13, 2025Updated last year
- MTU-Bench: A Multi-granularity Tool-Use Benchmark for Large Language Models☆60Jul 24, 2025Updated last year
- 基于大模型ChatGLM,微调方式为LORA,集SFT、RM、PPO算法为一体项目☆14Jun 20, 2023Updated 3 years ago
- 用于北航研究生考试自救☆24Jun 23, 2026Updated last month
- ☆13Jul 31, 2025Updated last year
- Benchmarking Language Agents Under Controllable and Extreme Context Growth☆51Apr 29, 2026Updated 3 months ago