[ICLR 2026 π₯] Dr.LLM: Dynamic Layer Routing in LLMs
β57Apr 24, 2026Updated 4 months ago
Alternatives and similar repositories for dr-llm
Users that are interested in dr-llm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-Agent LLM Evaluation Docs: https://maseval.readthedocs.io/β37Jul 5, 2026Updated last month
- [NAACL 2025 π₯] CAMEL-Bench is an Arabic benchmark for evaluating multimodal models across eight domains with 29,000 questions.β38Apr 17, 2025Updated last year
- β19Jul 24, 2023Updated 3 years ago
- transformer layers behavior as paintersπ§βπ¨β15May 6, 2025Updated last year
- Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Modelsβ19Jan 21, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- β17Jul 24, 2023Updated 3 years ago
- Code and data for "Timo: Towards Better Temporal Reasoning for Language Models" (COLM 2024)β26Oct 23, 2024Updated last year
- [CVPR 2025 π₯]A Large Multimodal Model for Pixel-Level Visual Grounding in Videosβ105Apr 14, 2025Updated last year
- Official Repository of Paper "Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs"β15Sep 25, 2025Updated 11 months ago
- Source code of "Leaky Thoughts: Large Reasoning Models Are Not Private Thinkers" EMNLP 2025β17Jan 12, 2026Updated 7 months ago
- VideoMathQA is a benchmark designed to evaluate mathematical reasoning in real-world educational videosβ24May 7, 2026Updated 3 months ago
- [EMNLP 2026 Main] Bottom-up Policy Optimization: Your Language Model Policy Secretly Contains Internal Policiesβ60Feb 6, 2026Updated 6 months ago
- β26Jul 24, 2026Updated last month
- Better coding experience for Flaskβ16Aug 11, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [NAACL 2022] "Learning to Win Lottery Tickets in BERT Transfer via Task-agnostic Mask Training", Yuanxin Liu, Fandong Meng, Zheng Lin, Peβ¦β15Oct 18, 2022Updated 3 years ago
- SAM Adaptation using SVDβ12Jul 13, 2025Updated last year
- [EMNLP 2025] TokenSkip: Controllable Chain-of-Thought Compression in LLMsβ226Nov 30, 2025Updated 9 months ago
- Official repo of dataset-decomposition paper [NeurIPS 2024]β21Jan 8, 2025Updated last year
- [CVPR'26 Demo] Mobile-O: Unified Multimodal Understanding and Generation on Mobile Deviceβ158Apr 13, 2026Updated 4 months ago
- [NAACL'25 π SAC Award] Official code for "Advancing MoE Efficiency: A Collaboration-Constrained Routing (C2R) Strategy for Better Expertβ¦β16Feb 4, 2025Updated last year
- Source code of "C-SEO Bench: Does Conversational SEO Work?" NeurIPS D&B 2025β20Sep 28, 2025Updated 11 months ago
- Source code and data for ADEPT: A DEbiasing PrompT Framework (AAAI-23).β15Dec 13, 2024Updated last year
- AIN - The First Arabic Inclusive Large Multimodal Model. It is a versatile bilingual LMM excelling in visual and contextual understandingβ¦β54Mar 13, 2025Updated last year
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Source code of "Calibrating Large Language Models Using Their Generations Only", ACL2024β22Nov 20, 2024Updated last year
- [ICLR 2025] When Attention Sink Emerges in Language Models: An Empirical View (Spotlight)β165Jul 8, 2025Updated last year
- β11May 9, 2023Updated 3 years ago
- β10Dec 28, 2023Updated 2 years ago
- Experiments for "A Closer Look at In-Context Learning under Distribution Shifts"β18May 29, 2023Updated 3 years ago
- [ACL 2025 π₯] A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understandingβ79Aug 10, 2026Updated 3 weeks ago
- Official code of the paper "VideoMolmo: Spatio-Temporal Grounding meets Pointing"β57Jul 5, 2025Updated last year
- β42Nov 9, 2023Updated 2 years ago
- Official repository for "Stylized Adversarial Training" (TPAMI 2022)β11Dec 30, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICLR 2024] Official code for the paper "LLM Blueprint: Enabling Text-to-Image Generation with Complex and Detailed Prompts"β85May 18, 2024Updated 2 years ago
- Code for COLM 2026 Paper "Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs"β30Jul 1, 2026Updated last month
- [SIGIR '25] This is the code repo for our SIGIR '25 paper: Enhancing the Patent Matching Capability of Large Language Models via Memory Gβ¦β19Apr 22, 2025Updated last year
- ImageNet-12k subset of ImageNet-21k (fall11)β23Jun 13, 2023Updated 3 years ago
- β45Jan 30, 2026Updated 7 months ago
- β23Dec 17, 2024Updated last year
- Official code for the paper Towards Fully Exploiting LLM Internal States to Enhance Knowledge Boundary Perception. The code is based on tβ¦β21Aug 5, 2025Updated last year