[ICLR 2026 π₯] Dr.LLM: Dynamic Layer Routing in LLMs
β57Apr 24, 2026Updated 4 months ago
Alternatives and similar repositories for dr-llm
Users that are interested in dr-llm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-Agent LLM Evaluation Docs: https://maseval.readthedocs.io/β38Jul 5, 2026Updated 2 months ago
- The official implementation of HybridNorm: Towards Stable and Efficient Transformer Training via Hybrid Normalizationβ19Mar 7, 2025Updated last year
- transformer layers behavior as paintersπ§βπ¨β15May 6, 2025Updated last year
- [ICCV2025] Hierarchical Visual Prompt Learning for Continual Video Instance Segmentationβ14Feb 18, 2026Updated 7 months ago
- Video-CoM: Interactive Video Reasoning via Chain of Manipulationsβ23Sep 5, 2026Updated 2 weeks ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Modelsβ19Sep 5, 2026Updated 2 weeks ago
- β17Jul 24, 2023Updated 3 years ago
- Code and data for "Timo: Towards Better Temporal Reasoning for Language Models" (COLM 2024)β26Oct 23, 2024Updated last year
- [CVPR 2025 π₯]A Large Multimodal Model for Pixel-Level Visual Grounding in Videosβ105Sep 5, 2026Updated 2 weeks ago
- Official Repository of Paper "Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs"β15Sep 25, 2025Updated 11 months ago
- Official repository of paper titled "D3Former: Debiased Dual Distilled Transformer for Incremental Learning".β25Jul 10, 2023Updated 3 years ago
- VideoMathQA is a benchmark designed to evaluate mathematical reasoning in real-world educational videosβ25Sep 5, 2026Updated 2 weeks ago
- xRouter: Training Cost-Aware LLMs Orchestration System via Reinforcement Learningβ34Jun 2, 2026Updated 3 months ago
- [EMNLP 2026 Main] Bottom-up Policy Optimization: Your Language Model Policy Secretly Contains Internal Policiesβ60Feb 6, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [NAACL 2022] "Learning to Win Lottery Tickets in BERT Transfer via Task-agnostic Mask Training", Yuanxin Liu, Fandong Meng, Zheng Lin, Peβ¦β15Oct 18, 2022Updated 3 years ago
- [EMNLP 2025] TokenSkip: Controllable Chain-of-Thought Compression in LLMsβ226Nov 30, 2025Updated 9 months ago
- Official repo of dataset-decomposition paper [NeurIPS 2024]β21Sep 11, 2026Updated last week
- [NAACL'25 π SAC Award] Official code for "Advancing MoE Efficiency: A Collaboration-Constrained Routing (C2R) Strategy for Better Expertβ¦β17Feb 4, 2025Updated last year
- Source code of "C-SEO Bench: Does Conversational SEO Work?" NeurIPS D&B 2025β20Sep 28, 2025Updated 11 months ago
- Source code and data for ADEPT: A DEbiasing PrompT Framework (AAAI-23).β15Dec 13, 2024Updated last year
- [ICLR 2025] When Attention Sink Emerges in Language Models: An Empirical View (Spotlight)β165Jul 8, 2025Updated last year
- β11May 9, 2023Updated 3 years ago
- ABB 140 Robot Draws a Given Pictureβ14Oct 17, 2020Updated 5 years ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Experiments for "A Closer Look at In-Context Learning under Distribution Shifts"β18May 29, 2023Updated 3 years ago
- [ACL 2025 π₯] A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understandingβ79Aug 10, 2026Updated last month
- Official code of the paper "VideoMolmo: Spatio-Temporal Grounding meets Pointing"β57Jul 5, 2025Updated last year
- Official repository for "Stylized Adversarial Training" (TPAMI 2022)β11Dec 30, 2022Updated 3 years ago
- [ICLR 2024] Official code for the paper "LLM Blueprint: Enabling Text-to-Image Generation with Complex and Detailed Prompts"β85May 18, 2024Updated 2 years ago
- β14Nov 26, 2021Updated 4 years ago
- [EMNLP'23] ClimateGPT: a specialized LLM for conversations related to Climate Change and Sustainability topics in both English and Arabiβ¦β80Sep 24, 2024Updated last year
- [SIGIR '25] This is the code repo for our SIGIR '25 paper: Enhancing the Patent Matching Capability of Large Language Models via Memory Gβ¦β19Apr 22, 2025Updated last year
- LaTeXDataHub is an open-source platform dedicated to the sharing and contribution of real-world LaTeX image datasets and their annotationβ¦β12Aug 13, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- β47Jan 30, 2026Updated 7 months ago
- β23Dec 17, 2024Updated last year
- β37Jul 24, 2023Updated 3 years ago
- Official code for the paper Towards Fully Exploiting LLM Internal States to Enhance Knowledge Boundary Perception. The code is based on tβ¦β21Aug 5, 2025Updated last year
- The official data and code for EMNLP 2023 main conference paper: CRT-QA: A Dataset of Complex Reasoning Question Answering over Tabular Dβ¦β13May 19, 2025Updated last year
- Official code release for Delta Activations: A Representation for Finetuned Large Language Modelsβ21Sep 5, 2025Updated last year
- LLM play 20questions with itselfβ13Mar 31, 2023Updated 3 years ago