SWE-Lego: Pushing the Limits of Supervised Fine-tuning for Software Issue Resolving
☆74Feb 28, 2026Updated 6 months ago
Alternatives and similar repositories for SWE-Lego
Users that are interested in SWE-Lego are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Jan 14, 2026Updated 7 months ago
- ☆50Mar 6, 2026Updated 5 months ago
- Advances and Frontiers of LLM-based Issue Resolution in Software Engineering A Comprehensive Survey☆87Aug 12, 2026Updated 3 weeks ago
- Learning MLPs to replace GNN☆10Jun 3, 2023Updated 3 years ago
- PhyX: Does Your Model Have the "Wits" for Physical Reasoning?☆55Mar 16, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The Github repo for our survey paper: "Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large…☆154Apr 15, 2026Updated 4 months ago
- ☆18May 18, 2025Updated last year
- open source SWE-Atlas☆70Aug 20, 2026Updated last week
- [ASE 2025] CoSIL: Issue Localization via Iteritive Code Graph Searching☆25May 31, 2026Updated 3 months ago
- 🚀 First survey on Attention Sink in Transformers — 200+ papers on utilization, interpretation, and mitigation.☆141Jun 5, 2026Updated 2 months ago
- [COLM 2025] Official repository for R2E-Gym: Procedural Environment Generation and Hybrid Verifiers for Scaling Open-Weights SWE Agents☆328Jul 13, 2025Updated last year
- Code for paper ”Language Versatilists vs. Specialists: An Empirical Revisiting on Multilingual Transfer Ability“☆15Jun 13, 2023Updated 3 years ago
- ☆23Oct 22, 2024Updated last year
- High-performance C++ inference engine for Diffusion Language Models (LLaDA, SEDD, MDLM)☆17Apr 5, 2026Updated 4 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆44Oct 28, 2025Updated 10 months ago
- [FSE'2026] SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarks☆191May 12, 2026Updated 3 months ago
- ☆18Mar 28, 2026Updated 5 months ago
- ☆13Aug 9, 2023Updated 3 years ago
- Code for paper: "Executing Arithmetic: Fine-Tuning Large Language Models as Turing Machines"☆10Oct 11, 2024Updated last year
- Code for M4LE: A Multi-Ability Multi-Range Multi-Task Multi-Domain Long-Context Evaluation Benchmark for Large Language Models☆23Jul 27, 2024Updated 2 years ago
- ☆10Apr 15, 2023Updated 3 years ago
- ☆11Dec 8, 2024Updated last year
- SWE-Exp: Experience-Driven Software Issue Resolution☆45Oct 17, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A benchmark for coding agents☆39Jun 25, 2026Updated 2 months ago
- Reproducing R1 for Code with Reliable Rewards☆13Apr 9, 2025Updated last year
- NVFP4 Flash-Attention 4 on BlackWell☆53Jul 23, 2026Updated last month
- The official code of "Towards Long-horizon Agentic Multimodal Search"☆29Apr 17, 2026Updated 4 months ago
- This is an official code for the paper: TestExplora: Benchmarking LLMs for Proactive Bug Discovery via Repository-Level Test Generation☆28Mar 26, 2026Updated 5 months ago
- Dream-VL and Dream-VLA, a diffusion VLM and a diffusion VLA.☆114Jan 14, 2026Updated 7 months ago
- Flash Attention implementation that returns both output and attention scores. High-performance, memory-efficient attention with score ext…☆16Feb 6, 2026Updated 6 months ago
- [COLM 2026] Official implementation for "MonitorBench: A Comprehensive Benchmark for Chain-of-Thought Monitorability in Large Language Mo…☆20Apr 23, 2026Updated 4 months ago
- The official implement of "Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings"☆18Dec 5, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 6 months ago
- [ICLR2026🔥Oral] SwingArena: Competitive Programming Arena for Long-context GitHub Issue Solving☆15Feb 26, 2026Updated 6 months ago
- Evals Harness for $OneMillion-Bench☆50Jul 30, 2026Updated last month
- ☆63Jul 1, 2026Updated 2 months ago
- Source code for Truth-Aware Context Selection: Mitigating the Hallucinations of Large Language Models Being Misled by Untruthful Contexts☆17Sep 2, 2024Updated 2 years ago
- ☆15Nov 12, 2025Updated 9 months ago
- Benchmark Test-Time Scaling of General LLM Agents☆23Apr 14, 2026Updated 4 months ago