The theory of mind module for the SWE agent
☆115May 28, 2026Updated last month
Alternatives and similar repositories for ToM-SWE
Users that are interested in ToM-SWE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆48Jun 11, 2026Updated last month
- ☆14Jul 5, 2024Updated 2 years ago
- Dataset collection and training code for "Ask Your Humans: Using Human Instructions to Improve Generalization in Reinforcement Learning"☆11Apr 8, 2025Updated last year
- Easiest way to build custom agents, in a no-code notion style editor, using simple macros.☆34Nov 8, 2024Updated last year
- ☆14Oct 8, 2025Updated 9 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- LLMatic is a 2-archive QD algorithm that uses LLMs to mutate the networks. Tested for Neural Architecture search but can easily be used f…☆21Aug 14, 2024Updated last year
- New testbed of interactive SWE tasks for coding agents, set in a realistic multi-turn developer driven environment☆24Jun 30, 2026Updated 3 weeks ago
- ☆31Apr 7, 2026Updated 3 months ago
- ☆19Dec 29, 2025Updated 6 months ago
- A clean, modular SDK for building AI agents with OpenHands V1.☆926Updated this week
- Universal REPL Bridge for LLMs - Tab completion, interactive prompts, TUI support☆20Nov 24, 2025Updated 8 months ago
- Agent computer interface for AI software engineer.☆133Apr 16, 2026Updated 3 months ago
- ⚔️ OpenHands PR Arena ⚔️ is a platform for evaluating and benchmarking agentic coding assistants through paired pull request (PR) generat…☆17Dec 15, 2025Updated 7 months ago
- Animated thinking/working/tool indicators for pi coding agent — 21 terminal animations with ANSI true color and Nerd Font glyphs☆20Jun 8, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- SWE-Together: Evaluating Coding Agents in Interactive User Sessions☆50Jul 7, 2026Updated 2 weeks ago
- ☆33Mar 7, 2026Updated 4 months ago
- ☆33Oct 15, 2025Updated 9 months ago
- Provides a minimal implementation to extract FLAN datasets for further processing☆11Feb 1, 2023Updated 3 years ago
- An overengineered, open-source task queue for Google Jules power users☆15Feb 16, 2026Updated 5 months ago
- ☆50May 23, 2026Updated 2 months ago
- Continual harness optimization☆66May 14, 2026Updated 2 months ago
- Public registry for OpenHands extensions.☆128Updated this week
- Official code for "How2Everything: Mining the Web for How-To Procedures to Evaluate and Improve LLMs"☆24Feb 10, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Hub repo for OpenWork agents, skills, and commands☆25Mar 27, 2026Updated 3 months ago
- Marketplace ML experiment - training without backprop☆28Sep 9, 2025Updated 10 months ago
- ☆15Feb 25, 2026Updated 5 months ago
- ☆39Aug 4, 2025Updated 11 months ago
- TUI conversation explorer for Claude Code & OpenCode☆20Aug 21, 2025Updated 11 months ago
- Small, simple agent task environments for training and evaluation☆20Nov 1, 2024Updated last year
- The web API server that runs program codes in an isolated environment using Docker.☆18Jul 20, 2023Updated 3 years ago
- ☆19Dec 20, 2025Updated 7 months ago
- The project page for "SCITAB: A Challenging Benchmark for Compositional Reasoning and Claim Verification on Scientific Tables"☆23Dec 21, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Demo code for "The Agent Is a Workflow That Writes Itself"☆22May 14, 2026Updated 2 months ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 3 months ago
- ☆15Feb 10, 2026Updated 5 months ago
- quickly mix LLMs together☆19Jan 5, 2025Updated last year
- [ICML 2026] SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark☆22May 6, 2026Updated 2 months ago
- ☆13Oct 18, 2023Updated 2 years ago
- SDK-first orchestration platform for managing AI coding agent fleets: BYOA key rotation, real-time WebSocket dashboard, DCG integration, …☆23Updated this week