New state-of-the-art bounds for open problems
☆125Apr 12, 2026Updated 4 months ago
Alternatives and similar repositories for EinsteinArena-new-SOTA
Users that are interested in EinsteinArena-new-SOTA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆51Updated this week
- ☆31Mar 13, 2026Updated 5 months ago
- ☆45Nov 6, 2025Updated 9 months ago
- ☆20Dec 8, 2025Updated 8 months ago
- ☆12May 30, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- An ARC-AGI solution using Agentica from Symbolica☆312Feb 12, 2026Updated 6 months ago
- Outputs from the Deep Writer☆16Sep 11, 2024Updated last year
- 🌌 Towards a Digital Pluriverse: “a world where many worlds may fit”☆15Feb 20, 2022Updated 4 years ago
- ☆21Jan 14, 2026Updated 7 months ago
- A transformer that executes a one-instruction Turing-complete computer — two approaches: hand-coded weights (no training) and learned fro…☆41Mar 3, 2026Updated 6 months ago
- FrontierSmith, a new system that uses AI to synthesize open-ended coding problems at scale☆52May 30, 2026Updated 3 months ago
- Programmatic memory for long-horizon LLM agents: the harness appends everything to one log, and the agent searches it with code. 97.4% on…☆420Aug 21, 2026Updated 2 weeks ago
- Official code, models, and dataset for "Evolution Fine-Tuning (EFT): Learning to Discover Across 371 Optimization Tasks"☆28Jun 30, 2026Updated 2 months ago
- [EMNLP 2024] Tree of Problems: Improving structured problem solving with compositionality☆20Mar 4, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP26 Findings] Official repository for DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning☆27Feb 12, 2026Updated 6 months ago
- [ICML 2026] Code for V1: Unifying Generation and Self-Verification for Parallel Reasoners.☆39Mar 5, 2026Updated 5 months ago
- [ICLR 26] The official code repository for the paper "Mirage or Method? How Model–Task Alignment Induces Divergent RL Conclusions".☆19Feb 9, 2026Updated 6 months ago
- Official code and dataset for our paper: RefineBench: Evaluating Refinement Capability of Language Models via Checklists☆17Dec 1, 2025Updated 9 months ago
- ☆11Aug 10, 2024Updated 2 years ago
- Forecasting scientific progress with AI☆31Aug 16, 2026Updated 2 weeks ago
- ☆634May 24, 2026Updated 3 months ago
- Code repository for the ICML 2026 Oral paper "Characterizing, Evaluating, and Optimizing Complex Reasoning".☆20Jun 21, 2026Updated 2 months ago
- The official repository of "Position: Agentic Evolution is the Path to Evolving LLMs".☆771Aug 22, 2026Updated last week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- AI-Driven Scientific, Algorithmic, and Systems Discovery☆628Updated this week
- Retrieval is CheapShow Me the Code: Executable Multi-Hop Reasoning for Retrieval-Augmented Generation☆26May 14, 2026Updated 3 months ago
- A web app that uses machine learning to recommend the most suitable journals based on the text content of your preprint☆22May 23, 2023Updated 3 years ago
- [ICML 2026] Official implementation for paper: Learning Self-Correction in Vision–Language Models via Rollout Augmentation☆16Jun 4, 2026Updated 3 months ago
- Chaos engineering for AI agents☆28Jan 2, 2026Updated 8 months ago
- A complete cognitive architecture on S³. One Hamilton product runs memory, learning, attention, and affect☆27Jun 1, 2026Updated 3 months ago
- ☆56Mar 13, 2026Updated 5 months ago
- VizHub Platform☆16Sep 4, 2025Updated last year
- Example for a Monty-enabled RLM in DSPy☆20Feb 16, 2026Updated 6 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Link Python logic with Svelte interfaces for simple demos☆14Jan 9, 2025Updated last year
- ☆79May 31, 2026Updated 3 months ago
- Document layer for agent work☆30Jul 27, 2026Updated last month
- ☆22May 10, 2026Updated 3 months ago
- ☆27Mar 7, 2023Updated 3 years ago
- Staging area for a public release of Theorizer☆172May 2, 2026Updated 4 months ago
- ALMA (Automated meta-Learning of Memory designs for Agentic systems) is a framework that meta-learns memory designs to replace human-engi…☆292Apr 8, 2026Updated 4 months ago