The repo for SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single Pass
☆97May 23, 2026Updated 2 months ago
Alternatives and similar repositories for SHINE
Users that are interested in SHINE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official implementation of LIFT: Improving Long Context Understanding of Large Language Models through Long Input Fine-Tuning☆15Mar 14, 2025Updated last year
- Official repository for the paper Number Cookbook: Number Understanding of Language Models and How to Improve It.☆22Mar 31, 2025Updated last year
- Hypernetworks that update LLMs to remember factual information☆794Jun 15, 2026Updated last month
- Design hardware-friendly model architectures and migrate existing LLMs with minimal performance loss☆497Jul 14, 2026Updated 3 weeks ago
- Towards a Mechanistic Understanding of Large Reasoning Models: A Survey of Training, Inference, and Failures☆34Jan 29, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- exploring whether LLMs perform case-based or rule-based reasoning☆31Mar 2, 2024Updated 2 years ago
- ☆21Mar 26, 2026Updated 4 months ago
- The official implementation of the paper "MLP Memory: A Retriever-Pretrained Memory for Large Language Models". (ICLR 2026)☆70Jun 11, 2026Updated last month
- SeRL: Self-Play Reinforcement Learning for Large Language Models with Limited Data☆24Jan 24, 2026Updated 6 months ago
- Repo of Paper: delta-Mem: Efficient Online Memory for Large Language Models☆51May 27, 2026Updated 2 months ago
- (ACM MM24) This is the offical repository of GIST: Improving Parameter Efficient Fine Tuning via Knowledge Interaction.☆11Jan 28, 2024Updated 2 years ago
- ☆17Jun 10, 2025Updated last year
- ☆80Feb 6, 2026Updated 6 months ago
- (ICME24) This is the offical repository of iDAT: inverse Distillation Adapter-Tuning.☆13Apr 3, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 电子科技大学软院期末复习汇总☆19Jun 4, 2023Updated 3 years ago
- Official code repository for the paper "ToMAP: Training Opponent-Aware LLM Persuaders with Theory of Mind"☆25Sep 25, 2025Updated 10 months ago
- PiSSA: Principal Singular Values and Singular Vectors Adaptation of Large Language Models(NeurIPS 2024 Spotlight)☆428Jun 30, 2025Updated last year
- [NAACL 2025] The official implementation of paper "Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language M…☆28Mar 14, 2024Updated 2 years ago
- Links to publications that focus on the interpretation and analysis of in-context learning☆14Oct 17, 2024Updated last year
- Metrics for evaluating biological sequence design☆15Jul 22, 2026Updated 2 weeks ago
- An in-context learning research testbed☆19Mar 16, 2025Updated last year
- Contrastive Dialogue Disentanglement via Clustering☆13Apr 26, 2023Updated 3 years ago
- [NeurIPS'24] HippoRAG is a novel RAG framework inspired by human long-term memory that enables LLMs to continuously integrate knowledge a…☆13Mar 6, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆17Apr 9, 2025Updated last year
- [WWW2024 Oral] Harnessing Multi-Role Capabilities of Large Language Models for Open-Domain Question Answering☆15Apr 22, 2025Updated last year
- ☆30Feb 27, 2026Updated 5 months ago
- Activation-Steered Compression☆17Jan 30, 2026Updated 6 months ago
- The official implementation for "ImageDoctor: Diagnosing Text-to-Image Generation via Grounded Image Reasoning"☆15Jul 31, 2026Updated last week
- Working with images in frequency space☆10Nov 5, 2020Updated 5 years ago
- ☆16Jun 4, 2025Updated last year
- Automatically Update LLM Papers Daily using Github Actions. Ref: https://github.com/Vincentqyw/cv-arxiv-daily☆10Aug 3, 2026Updated last week
- We introduce new approach, Token Reduction using CLIP Metric (TRIM), aimed at improving the efficiency of MLLMs without sacrificing their…☆22Jan 11, 2026Updated 6 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ASE2024] Mutual Learning-Based Framework for Enhancing Robustness of Code Models via Adversarial Training☆11Sep 13, 2024Updated last year
- The official repository for SkyLadder: Better and Faster Pretraining via Context Window Scheduling☆43Dec 29, 2025Updated 7 months ago
- Co-Supervised Learning: Improving Weak-to-Strong Generalization with Hierarchical Mixture of Experts☆15Feb 26, 2024Updated 2 years ago
- ☆18Oct 15, 2025Updated 9 months ago
- [EMNLP 2024] Quantize LLM to extremely low-bit, and finetune the quantized LLMs☆16Jul 18, 2024Updated 2 years ago
- On the Robustness of GUI Grounding Models Against Image Attacks☆12Apr 8, 2025Updated last year
- ☆14Oct 17, 2024Updated last year