Hypernetworks that update LLMs to remember factual information
☆794Jun 15, 2026Updated last month
Alternatives and similar repositories for doc-to-lora
Users that are interested in doc-to-lora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Hypernetworks that adapt LLMs for specific benchmark tasks using only textual task description as the input☆1,298Jun 8, 2025Updated last year
- The repo for SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single Pass☆97May 23, 2026Updated 2 months ago
- Code for Fast-weight Product Key Memory (FwPKM)☆20Mar 18, 2026Updated 4 months ago
- We propose a novel modular framework that learns to dynamically mix low-rank adapters (LoRAs) to improve visual analogy learning, enablin…☆75Aug 2, 2026Updated last week
- Official JAX implementation of End-to-End Test-Time Training for Long Context☆631Feb 15, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.☆5,387Updated this week
- [ICML2026] From Statics to Dynamics: Physics-Aware Image Editing with Latent Transition Priors☆93Apr 30, 2026Updated 3 months ago
- Cuda kernels for leveraging LLM sparsity to improve throughput and decrease the memory requirements during inference and training.☆256Jun 29, 2026Updated last month
- DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation☆244Feb 18, 2026Updated 5 months ago
- 🌋LavaSR: Fast Speech restoration and enhancement☆571Jun 19, 2026Updated last month
- [arxiv: 2503.23895] Dynamic Parametric Retrieval Augmented Generation for Test-time Knowledge Enhancement☆182Aug 14, 2025Updated 11 months ago
- Reinforcement Learning via Self-Distillation (SDPO)☆1,043Jul 1, 2026Updated last month
- Zero and Few-shot document level relation extraction / ⚠️ Development moved to: https://github.com/cea-list-lasti/glidre☆18Mar 13, 2026Updated 4 months ago
- Official PyTorch Implementation for Learning a Generative Meta-Model of LLM Activations, ICML 2026☆91Apr 30, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ThetaEvolve: Test-time Learning on Open Problems, enabling RL training on AlphaEvolve/OpenEvolve and emphasizing scaling test-time comput…☆173Feb 27, 2026Updated 5 months ago
- OpenClaw-RL: Train any agent simply by talking☆5,627May 23, 2026Updated 2 months ago
- [ICML 2026] ReCo: In-Context Generation with Regional Constraints for Instructional Video Editing☆175May 26, 2026Updated 2 months ago
- Optimize prompts, code, and more with AI-powered Reflective Optimization☆6,037Updated this week
- Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement…☆10,571Updated this week
- Internal Coherence Maximization (ICM): A Label-Free, Unsupervised Training Framework for LLMs☆27Sep 5, 2025Updated 11 months ago
- Optimizing inference proxy for LLMs☆4,235Jul 18, 2026Updated 3 weeks ago
- The Official PyTorch implementation of Shared LoRA Subspaces for almost Strict Continual Learning☆34May 7, 2026Updated 3 months ago
- ☆26Feb 10, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR'26] VecGlypher: Unified Vector Glyph Generation with Language Models☆139Feb 26, 2026Updated 5 months ago
- [CVPR 2026] Adaptive Spectral Feature Forecasting for Diffusion Sampling Acceleration☆129Apr 30, 2026Updated 3 months ago
- A Tree Search Library with Flexible API for LLM Inference-Time Scaling☆559Feb 5, 2026Updated 6 months ago
- Training teachers with reinforcement learning able to make LLMs learn how to reason for test time scaling.☆365Jun 23, 2025Updated last year
- OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis☆1,127Jun 10, 2026Updated 2 months ago
- Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models☆4,581Jan 14, 2026Updated 6 months ago
- 🖥 Neural Computers' Data Engine☆201May 19, 2026Updated 2 months ago
- DySCO: Dynamic Attention-Scaling Decoding for Long-Context LMs☆17May 30, 2026Updated 2 months ago
- ShinkaEvolve: Towards Open-Ended and Sample-Efficient Program Evolution 🧬☆1,325Jul 31, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Extending the Context of Pretrained LLMs by Dropping Their Positional Embedding☆220Jan 12, 2026Updated 6 months ago
- [CVPR 2026] ShowUI-π: Flow-based Generative Models as GUI Dexterous Hands☆132Apr 22, 2026Updated 3 months ago
- RePo: Language Models with Context Re-Positioning☆83Mar 30, 2026Updated 4 months ago
- Darwin Gödel Machine: Open-Ended Evolution of Self-Improving Agents☆2,215Aug 13, 2025Updated 11 months ago
- HY-WU (Part I): An Extensible Functional Neural Memory Framework and An Instantiation in Text-Guided Image Editing☆297Mar 18, 2026Updated 4 months ago
- ALMA (Automated meta-Learning of Memory designs for Agentic systems) is a framework that meta-learns memory designs to replace human-engi…☆251Apr 8, 2026Updated 4 months ago
- Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B☆1,550Jun 17, 2026Updated last month