SR²AM: Efficient Agentic Reasoning Through Self-Regulated Simulative Planning
☆21May 22, 2026Updated 2 months ago
Alternatives and similar repositories for sr2am
Users that are interested in sr2am are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆29Apr 11, 2026Updated 3 months ago
- [ICML 2026] Code for Equilibrium Reasoners: learning attractor dynamics for scalable reasoning☆45Jun 1, 2026Updated 2 months ago
- ☆16Feb 12, 2020Updated 6 years ago
- ☆14Dec 25, 2024Updated last year
- MemSearcher is a search agent that keeps a compact, iteratively-updated memory instead of the full interaction history, trained end-to-en…☆27Jun 29, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆23Apr 5, 2026Updated 3 months ago
- ☆35Nov 11, 2025Updated 8 months ago
- [ICML 2026] NITP: Next Implicit Token Prediction for LLM Pre-training☆34May 26, 2026Updated 2 months ago
- Implementation of paper: Extending and Analyzing Self-Supervised Learning Across Domains☆10Jan 10, 2021Updated 5 years ago
- An Ultra-Long Output Reinforcement Learning Approach☆23Jul 31, 2025Updated last year
- Code for EMNLP2023 paper "MolCA: Molecular Graph-Language Modeling with Cross-Modal Projector and Uni-Modal Adapter".☆12Dec 27, 2023Updated 2 years ago
- Loading MNIST handwritten digits with Octave☆12Oct 20, 2017Updated 8 years ago
- ☆71May 12, 2026Updated 2 months ago
- ☆14Mar 16, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Super Mario Bros. (NES) gameplay dataset for machine learning.☆13Jul 22, 2025Updated last year
- Algorithms for latent compaction☆261Apr 22, 2026Updated 3 months ago
- COS-PLAY: Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Game Play☆29Jul 11, 2026Updated 3 weeks ago
- Code for the paper "Data Feedback Loops: Model-driven Amplification of Dataset Biases"☆18Sep 9, 2022Updated 3 years ago
- Benchmarking long-horizon chain-of-thought reasoning.☆40Apr 20, 2026Updated 3 months ago
- This repository is the official implementation of the Hybrid Self-Attention NEAT algorithm. It contains the code to reproduce the results…☆15Jun 19, 2023Updated 3 years ago
- 📸 Code and Dataset for our ACL 2023 paper: "MPCHAT: Towards Multimodal Persona-Grounded Conversation"☆22Sep 5, 2023Updated 2 years ago
- ☆13Jun 7, 2020Updated 6 years ago
- VCCA Pytorch Implementation on MNIST dataset☆16Apr 10, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- MEDDxAgent: A Unified Modular Agent Framework for Explainable Automatic Differential Diagnosis☆23Jun 13, 2025Updated last year
- Socratic-Zero is a fully autonomous framework that generates high-quality training data for mathematical reasoning☆37Oct 26, 2025Updated 9 months ago
- 3MDBench: Medical Multimodal Multi-agent Dialogue Benchmark☆24Sep 23, 2025Updated 10 months ago
- Evaluation repository of wikipedia index with Dria☆10Mar 14, 2024Updated 2 years ago
- 🎉My Collections of CUDA Kernels~☆11Jun 25, 2024Updated 2 years ago
- Submodule of evalverse forked from [google-research/instruction_following_eval](https://github.com/google-research/google-research/tree/m…☆15May 4, 2024Updated 2 years ago
- MDPO: Overcoming the Training-Inference Divide of Masked Diffusion Language Models☆45Jan 28, 2026Updated 6 months ago
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 3 months ago
- [ACL 2024] Dataset and Code of "ImplicitAVE: An Open-Source Dataset and Multimodal LLMs Benchmark for Implicit Attribute Value Extraction…☆17Jun 10, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Curated LLM (ICML 2024)☆15Oct 23, 2024Updated last year
- MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism☆18Nov 18, 2025Updated 8 months ago
- ☆17May 21, 2020Updated 6 years ago
- Code of hi-GCN☆28Jun 26, 2024Updated 2 years ago
- ☆16Jun 15, 2026Updated last month
- All-in-one environment to use Dria, the collective knowledge for AI.☆14Mar 15, 2024Updated 2 years ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 3 months ago