Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents
☆31Apr 16, 2026Updated 5 months ago
Alternatives and similar repositories for MemoryTransferLearning
Users that are interested in MemoryTransferLearning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation for the paper "Video-Based Reward Modeling for Computer-Use Agents"☆17Mar 14, 2026Updated 6 months ago
- We revisit the Platonic Representation Hypothesis using calibrated representational similarity metrics with statistical guarantees.☆38Jun 24, 2026Updated 3 months ago
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models☆22Apr 14, 2026Updated 5 months ago
- ☆19May 25, 2026Updated 4 months ago
- Code repo for paper: Effective Strategies for Asynchronous Software Engineering Agents☆73Apr 2, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems (ICLR'26)☆32Nov 3, 2025Updated 11 months ago
- The offical repo for "Parallel-Probe: Towards Efficient Parallel Thinking via 2D Probing"☆21Feb 3, 2026Updated 8 months ago
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆23Jun 2, 2026Updated 4 months ago
- ☆17Updated this week
- In-Context Reinforcement Learning for Tool Use in Large Language Models☆48Mar 26, 2026Updated 6 months ago
- Official Implementation of MGS: Matryoshka Gaussian Splatting☆40Jun 11, 2026Updated 3 months ago
- Energy-based Hallucination detection.☆25Mar 3, 2026Updated 7 months ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 5 months ago
- [EACL 2026] PaperSearchQA. Data generation pipeline for QA over scientific papers, suitable for RL training search agents☆37Feb 4, 2026Updated 7 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- MEMO: Memory-Augmented Model Context Optimization for Robust Multi-Turn Multi-Agent LLM Games☆30May 10, 2026Updated 4 months ago
- [🏆ECCV'26] Official Repo for SlowBA: An efficiency backdoor attack towards VLM-based GUI agents☆20Sep 14, 2026Updated 2 weeks ago
- Open-source test harness for AI agents. Stress-test production agents with adversarial multi-turn scenarios in CI☆30Aug 14, 2026Updated last month
- [ICML 2026] Code for V1: Unifying Generation and Self-Verification for Parallel Reasoners.☆39Mar 5, 2026Updated 6 months ago
- [ICLR 2026] Official implementation of "Enhancing Multi-Image Understanding Through Delimiter Token Scaling"☆17Jul 10, 2026Updated 2 months ago
- [CVPR 2026 Oral] FINER: MLLMs Hallucinate under Fine-grained Negative Queries☆19Jul 6, 2026Updated 2 months ago
- A research framework for evaluating proactive AI assistants through active user simulation☆40May 23, 2026Updated 4 months ago
- ☆10Jun 15, 2024Updated 2 years ago
- GradMem: Learning to Write Context into Memory with Test-Time Gradient Descent [ICML 2026]☆41Sep 16, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [SIGIR 2026] "One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment"☆17Apr 21, 2026Updated 5 months ago
- PyTorch Implementation of LocAtViT in "Locality-Attending Vision Transformer" (ICLR 2026)☆20Aug 12, 2026Updated last month
- This is the codes of "DARE: Aligning LLM Agents with the R Statistical Ecosystem via Distribution-Aware Retrieval"☆15Aug 11, 2026Updated last month
- ☆18Mar 16, 2026Updated 6 months ago
- ☆38Apr 3, 2026Updated 6 months ago
- ☆20Jul 1, 2026Updated 3 months ago
- Group Evolving Agents: Open-Ended Self-Improvement via Experience Sharing☆134Apr 7, 2026Updated 5 months ago
- [ACL 2026 SAC Highlight Award] Can We Predict Before Executing Machine Learning Agents?☆24Jul 7, 2026Updated 2 months ago
- Official code of Geometric Autoencoder for Diffusion Models.☆22Mar 12, 2026Updated 6 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICML 2026] Official Implementation of Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diff…☆23Mar 4, 2026Updated 6 months ago
- ☆15Feb 21, 2024Updated 2 years ago
- 4DEquine: Disentangling Motion and Appearance for 4D Equine Reconstruction from Monocular Video (CVPR2026)☆19Apr 12, 2026Updated 5 months ago
- ProAct is a framework designed to enable Large Language Model (LLM) agents to perform accurate, multi-turn lookahead reasoning in interac…☆18Feb 11, 2026Updated 7 months ago
- Official project page and code repository for WiT, a pixel space diffusion☆17Sep 14, 2026Updated 2 weeks ago
- Codebase for DeepVision-103K☆22Feb 21, 2026Updated 7 months ago
- [ICML2026] Reproduce Kimi K1.5/K2 RL algorithm and rollout system☆21Apr 9, 2026Updated 5 months ago