☆18Mar 28, 2026Updated 5 months ago
Alternatives and similar repositories for SWE-Compass
Users that are interested in SWE-Compass are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆88Jun 19, 2026Updated 2 months ago
- SWE-Swiss: A Multi-Task Fine-Tuning and RL Recipe for High-Performance Issue Resolution☆105Sep 24, 2025Updated 11 months ago
- An implementation of online data mixing for the Pile dataset, based on the GPT-NeoX library.☆14Jan 9, 2024Updated 2 years ago
- [ACL 2024 Findings] The official repo for "ConceptMath: A Bilingual Concept-wise Benchmark for Measuring Mathematical Reasoning of Large …☆26May 29, 2024Updated 2 years ago
- Official repository for our paper "FullStack Bench: Evaluating LLMs as Full Stack Coders"☆121May 7, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆42Jul 15, 2025Updated last year
- Official Code Repository for [AutoScale📈: Scale-Aware Data Mixing for Pre-Training LLMs] Published as a conference paper at **COLM 2025*…☆14Aug 8, 2025Updated last year
- Official Code of MEnvAgent☆25Feb 3, 2026Updated 6 months ago
- [FSE'2026] SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarks☆190May 12, 2026Updated 3 months ago
- [ICML 2026] SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark☆23May 6, 2026Updated 3 months ago
- The Source Code for DR3-Eval☆40Aug 12, 2026Updated 2 weeks ago
- ☆53Oct 28, 2025Updated 10 months ago
- The code and data for the paper JiuZhang3.0☆49May 26, 2024Updated 2 years ago
- The Source Code for WebCompass☆25May 2, 2026Updated 3 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ACL25' Findings] SWE-Dev is an SWE agent with a scalable test case construction pipeline.☆65Jul 21, 2025Updated last year
- ☆51Aug 5, 2025Updated last year
- ☆24May 7, 2026Updated 3 months ago
- ☆167May 13, 2026Updated 3 months ago
- ☆139May 8, 2025Updated last year
- We implement an efficient mechanism for compressing large networks by {\em tensorizing\/} network layers: i.e. mapping layers on to high-…☆11Jul 10, 2018Updated 8 years ago
- Simple code for the tutorial on Polynomial Nets.☆13Jan 19, 2023Updated 3 years ago
- ☆20Apr 9, 2025Updated last year
- [ICML 2024] UGrid: An Efficient-And-Rigorous Neural Multigrid Solver for Linear PDEs☆12Aug 7, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- SWE-Flow: Synthesizing Software Engineering Data in a Test-Driven Manner☆40Jun 29, 2025Updated last year
- ☆48Dec 12, 2024Updated last year
- Multi-agent synthetic data generation pipeline capable of generating and validating long horizon terminal/coding tasks for RL training☆74Jul 28, 2025Updated last year
- SimpleVQA: Multimodal Factuality Evaluation for Multimodal Large Language Models☆15Feb 20, 2025Updated last year
- [ICML 2026 Oral] Agent-native Mid-training for Software Engineering☆77Jun 7, 2026Updated 2 months ago
- ✨ RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems - ICLR 2024☆212Aug 16, 2024Updated 2 years ago
- Qwen1.5大模型微调、基于PEFT框架LoRA微调,在数据集HC3-Chinese上实现文本分类。☆12Jun 29, 2024Updated 2 years ago
- Heuristic filtering framework for RefineCode☆88Mar 13, 2025Updated last year
- [ISSTA'25] A GitHub issue resolution benchmark with multi-aspect diversity in programming languages, repository domains and modality of i…☆17Jun 13, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ACL 2025] Graph Aligned Large Language Models for Improved Source Code Understanding☆44May 18, 2025Updated last year
- MTU-Bench: A Multi-granularity Tool-Use Benchmark for Large Language Models☆60Jul 24, 2025Updated last year
- OpenClaw-RL: Personalize openclaw simply by talking to it☆16Feb 26, 2026Updated 6 months ago
- ☆14Jul 31, 2025Updated last year
- Benchmarking Language Agents Under Controllable and Extreme Context Growth☆56Apr 29, 2026Updated 4 months ago
- English and Chinese LaTeX template for reports/projects/proposal at Beijing Institute of Technology☆10Nov 19, 2020Updated 5 years ago
- ☆52Mar 9, 2026Updated 5 months ago