Hypernetworks that update LLMs to remember factual information
☆805Jun 15, 2026Updated 2 months ago
Alternatives and similar repositories for doc-to-lora
Users that are interested in doc-to-lora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Hypernetworks that adapt LLMs for specific benchmark tasks using only textual task description as the input☆1,300Jun 8, 2025Updated last year
- The repo for SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single Pass☆110Aug 18, 2026Updated last week
- Code for Fast-weight Product Key Memory (FwPKM)☆22Mar 18, 2026Updated 5 months ago
- We propose a novel modular framework that learns to dynamically mix low-rank adapters (LoRAs) to improve visual analogy learning, enablin…☆75Aug 2, 2026Updated 3 weeks ago
- Official JAX implementation of End-to-End Test-Time Training for Long Context☆682Feb 15, 2026Updated 6 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.☆5,559Updated this week
- [ICML2026] From Statics to Dynamics: Physics-Aware Image Editing with Latent Transition Priors☆93Apr 30, 2026Updated 4 months ago
- Cuda kernels for leveraging LLM sparsity to improve throughput and decrease the memory requirements during inference and training.☆257Jun 29, 2026Updated 2 months ago
- DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation☆306Feb 18, 2026Updated 6 months ago
- 🌋LavaSR: Fast Speech restoration and enhancement☆580Jun 19, 2026Updated 2 months ago
- [arxiv: 2503.23895] Dynamic Parametric Retrieval Augmented Generation for Test-time Knowledge Enhancement☆184Aug 14, 2025Updated last year
- Reinforcement Learning via Self-Distillation (SDPO)☆1,081Jul 1, 2026Updated last month
- Zero and Few-shot document level relation extraction / ⚠️ Development moved to: https://github.com/cea-list-lasti/glidre☆19Mar 13, 2026Updated 5 months ago
- Official PyTorch Implementation for Learning a Generative Meta-Model of LLM Activations, ICML 2026☆94Apr 30, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ThetaEvolve: Test-time Learning on Open Problems, enabling RL training on AlphaEvolve/OpenEvolve and emphasizing scaling test-time comput…☆176Feb 27, 2026Updated 6 months ago
- OpenClaw-RL: Train any agent simply by talking☆5,660May 23, 2026Updated 3 months ago
- [ICML 2026] ReCo: In-Context Generation with Regional Constraints for Instructional Video Editing☆176Aug 20, 2026Updated last week
- Optimize prompts, code, and more with AI-powered Reflective Optimization☆6,295Updated this week
- Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement…☆10,682Updated this week
- Internal Coherence Maximization (ICM): A Label-Free, Unsupervised Training Framework for LLMs☆27Sep 5, 2025Updated 11 months ago
- Optimizing inference proxy for LLMs☆4,257Jul 18, 2026Updated last month
- The Official PyTorch implementation of Shared LoRA Subspaces for almost Strict Continual Learning☆35May 7, 2026Updated 3 months ago
- ☆27Feb 10, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR'26] VecGlypher: Unified Vector Glyph Generation with Language Models☆142Feb 26, 2026Updated 6 months ago
- [CVPR 2026] Adaptive Spectral Feature Forecasting for Diffusion Sampling Acceleration☆134Apr 30, 2026Updated 4 months ago
- A Tree Search Library with Flexible API for LLM Inference-Time Scaling☆563Feb 5, 2026Updated 6 months ago
- OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis☆1,211Jun 10, 2026Updated 2 months ago
- Training teachers with reinforcement learning able to make LLMs learn how to reason for test time scaling.☆367Jun 23, 2025Updated last year
- Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models☆4,623Jan 14, 2026Updated 7 months ago
- 🖥 Neural Computers' Data Engine☆203May 19, 2026Updated 3 months ago
- ShinkaEvolve: Towards Open-Ended and Sample-Efficient Program Evolution 🧬☆1,361Aug 21, 2026Updated last week
- Extending the Context of Pretrained LLMs by Dropping Their Positional Embedding☆221Jan 12, 2026Updated 7 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [CVPR 2026] ShowUI-π: Flow-based Generative Models as GUI Dexterous Hands☆135Apr 22, 2026Updated 4 months ago
- RePo: Language Models with Context Re-Positioning☆84Mar 30, 2026Updated 5 months ago
- Darwin Gödel Machine: Open-Ended Evolution of Self-Improving Agents☆2,261Aug 13, 2025Updated last year
- HY-WU (Part I): An Extensible Functional Neural Memory Framework and An Instantiation in Text-Guided Image Editing☆394Mar 18, 2026Updated 5 months ago
- ALMA (Automated meta-Learning of Memory designs for Agentic systems) is a framework that meta-learns memory designs to replace human-engi…☆291Apr 8, 2026Updated 4 months ago
- Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B☆1,564Aug 14, 2026Updated 2 weeks ago
- Official repository for DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research☆700Jun 17, 2026Updated 2 months ago