Hypernetworks that update LLMs to remember factual information
☆796Jun 15, 2026Updated last month
Alternatives and similar repositories for doc-to-lora
Users that are interested in doc-to-lora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Hypernetworks that adapt LLMs for specific benchmark tasks using only textual task description as the input☆1,299Jun 8, 2025Updated last year
- The repo for SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single Pass☆97May 23, 2026Updated 2 months ago
- Code for Fast-weight Product Key Memory (FwPKM)☆20Mar 18, 2026Updated 4 months ago
- We propose a novel modular framework that learns to dynamically mix low-rank adapters (LoRAs) to improve visual analogy learning, enablin…☆75Aug 2, 2026Updated last week
- Official JAX implementation of End-to-End Test-Time Training for Long Context☆631Feb 15, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.☆5,404Updated this week
- [ICML2026] From Statics to Dynamics: Physics-Aware Image Editing with Latent Transition Priors☆93Apr 30, 2026Updated 3 months ago
- Cuda kernels for leveraging LLM sparsity to improve throughput and decrease the memory requirements during inference and training.☆256Jun 29, 2026Updated last month
- DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation☆243Feb 18, 2026Updated 5 months ago
- 🌋LavaSR: Fast Speech restoration and enhancement☆571Jun 19, 2026Updated last month
- [arxiv: 2503.23895] Dynamic Parametric Retrieval Augmented Generation for Test-time Knowledge Enhancement☆182Aug 14, 2025Updated 11 months ago
- Reinforcement Learning via Self-Distillation (SDPO)☆1,048Jul 1, 2026Updated last month
- Official PyTorch Implementation for Learning a Generative Meta-Model of LLM Activations, ICML 2026☆91Apr 30, 2026Updated 3 months ago
- ThetaEvolve: Test-time Learning on Open Problems, enabling RL training on AlphaEvolve/OpenEvolve and emphasizing scaling test-time comput…☆173Feb 27, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- OpenClaw-RL: Train any agent simply by talking☆5,626May 23, 2026Updated 2 months ago
- [ICML 2026] ReCo: In-Context Generation with Regional Constraints for Instructional Video Editing☆175May 26, 2026Updated 2 months ago
- Optimize prompts, code, and more with AI-powered Reflective Optimization☆6,056Updated this week
- Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement…☆10,577Updated this week
- Internal Coherence Maximization (ICM): A Label-Free, Unsupervised Training Framework for LLMs☆27Sep 5, 2025Updated 11 months ago
- Optimizing inference proxy for LLMs☆4,236Jul 18, 2026Updated 3 weeks ago
- The Official PyTorch implementation of Shared LoRA Subspaces for almost Strict Continual Learning☆34May 7, 2026Updated 3 months ago
- ☆26Feb 10, 2026Updated 6 months ago
- [CVPR'26] VecGlypher: Unified Vector Glyph Generation with Language Models☆139Feb 26, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [CVPR 2026] Adaptive Spectral Feature Forecasting for Diffusion Sampling Acceleration☆130Apr 30, 2026Updated 3 months ago
- A Tree Search Library with Flexible API for LLM Inference-Time Scaling☆559Feb 5, 2026Updated 6 months ago
- Training teachers with reinforcement learning able to make LLMs learn how to reason for test time scaling.☆365Jun 23, 2025Updated last year
- OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis☆1,129Jun 10, 2026Updated 2 months ago
- Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models☆4,582Jan 14, 2026Updated 6 months ago
- DySCO: Dynamic Attention-Scaling Decoding for Long-Context LMs☆17May 30, 2026Updated 2 months ago
- 🖥 Neural Computers' Data Engine☆201May 19, 2026Updated 2 months ago
- ShinkaEvolve: Towards Open-Ended and Sample-Efficient Program Evolution 🧬☆1,330Jul 31, 2026Updated last week
- Extending the Context of Pretrained LLMs by Dropping Their Positional Embedding☆220Jan 12, 2026Updated 6 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [CVPR 2026] ShowUI-π: Flow-based Generative Models as GUI Dexterous Hands☆132Apr 22, 2026Updated 3 months ago
- RePo: Language Models with Context Re-Positioning☆83Mar 30, 2026Updated 4 months ago
- Darwin Gödel Machine: Open-Ended Evolution of Self-Improving Agents☆2,215Aug 13, 2025Updated 11 months ago
- HY-WU (Part I): An Extensible Functional Neural Memory Framework and An Instantiation in Text-Guided Image Editing☆297Mar 18, 2026Updated 4 months ago
- ALMA (Automated meta-Learning of Memory designs for Agentic systems) is a framework that meta-learns memory designs to replace human-engi…☆251Apr 8, 2026Updated 4 months ago
- Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B☆1,550Jun 17, 2026Updated last month
- Official repository for DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research☆693Jun 17, 2026Updated last month