The original Shared Recurrent Memory Transformer implementation
☆39Aug 24, 2026Updated 3 weeks ago
Alternatives and similar repositories for srmt
Users that are interested in srmt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Cramming 1568 Tokens into a Single Vector and Back Again: Exploring the Limits of Embedding Space Capacity (ACL 2025, oral)☆35Jun 14, 2025Updated last year
- Natural Language Reinforcement Learning☆104Jul 30, 2025Updated last year
- Vintix: Action Model via In-Context Reinforcement Learning - - — ICML 2025☆51May 23, 2025Updated last year
- [ICML 2026] Official Implementation of Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diff…☆23Mar 4, 2026Updated 6 months ago
- [EMNLP26 Findings] Official repository for DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning☆27Feb 12, 2026Updated 7 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆11May 18, 2025Updated last year
- https://x.com/BlinkDL_AI/status/1884768989743882276☆28May 4, 2025Updated last year
- Collection of LLM completions for reasoning-gym task datasets☆31Jul 4, 2025Updated last year
- entropix style sampling + GUI☆27Oct 30, 2024Updated last year
- [ACL 2025] Agentic Reward Modeling: Integrating Human Preferences with Verifiable Correctness Signals for Reliable Reward Systems☆134Jun 11, 2025Updated last year
- Repository containing lectures from 2023 Machine Learning course☆12Mar 14, 2023Updated 3 years ago
- [IROS-2025] MAPF-GPT-DDG is a scalable decentralized multi-agent pathfinding (MAPF) solver based on imitation learning. It builds upon MA…☆68Feb 21, 2026Updated 7 months ago
- ☆15May 27, 2025Updated last year
- ☆25Oct 28, 2025Updated 10 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- THOUGHTSCULPT, a general reasoning and search method for complex tasks☆13Dec 13, 2024Updated last year
- The official implementation of the paper "A Dual-Space Framework for General Knowledge Distillation of Large Language Models".☆17Jan 4, 2026Updated 8 months ago
- Mitigating Partial Observability in Sequential Decision Processes via the Lambda Discrepancy☆24Oct 28, 2024Updated last year
- ☆24Jun 16, 2026Updated 3 months ago
- 动手学ROS2系列教程☆16Aug 26, 2021Updated 5 years ago
- ☆96Dec 6, 2024Updated last year
- Systematic evaluation framework that automatically rates overthinking behavior in large language models.☆103May 16, 2025Updated last year
- Code for paper: Unified Text-to-Image Generation and Retrieval☆15Jul 19, 2026Updated 2 months ago
- Official implementation of Self-Taught Agentic Long Context Understanding (ACL 2025).☆14Sep 22, 2025Updated 11 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The official GitHub page for the survey paper "A Survey of RWKV".☆33Jan 7, 2025Updated last year
- [EMNLP 2025🔥] UNComp: Can Matrix Entropy Uncover Sparsity? -- A Compressor Design from an Uncertainty-Aware Perspective☆20Jan 7, 2026Updated 8 months ago
- ☆15Apr 11, 2024Updated 2 years ago
- 🧠Plan-and-Budget: Training-free test-time reasoning framework for adaptive token allocation in large language models (ICLR 2026).☆17Mar 2, 2026Updated 6 months ago
- [AAMAS 2024] HiMAP: Learning Heuristics-Informed Policies for Large-Scale Multi-Agent Pathfinding☆14Mar 12, 2024Updated 2 years ago
- (ACL2025 oral) SCOPE: Optimizing KV Cache Compression in Long-context Generation☆36May 28, 2025Updated last year
- Resources for our paper: "Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training"☆175Oct 20, 2025Updated 11 months ago
- ☆47May 27, 2025Updated last year
- Source code for Pathfinding in Stochastic Environments paper.☆15Oct 27, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Benchmark Test-Time Scaling of General LLM Agents☆25Apr 14, 2026Updated 5 months ago
- [ICLR 2025] Large (Vision) Language Models are Unsupervised In-Context Learners☆22Jun 6, 2025Updated last year
- ☆33Jun 5, 2025Updated last year
- [ACL 2026 Main] Official Repo for Paper "Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Ali…☆16Jul 1, 2026Updated 2 months ago
- Repository for "Training Language Models To Explain Their Own Computations"☆37Jul 7, 2026Updated 2 months ago
- ☆16Jul 23, 2024Updated 2 years ago
- ☆18Feb 22, 2025Updated last year