Codebase for Instruction Following without Instruction Tuning
β36Sep 24, 2024Updated 2 years ago
Alternatives and similar repositories for implicit-ins
Users that are interested in implicit-ins are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Code Repository for [AutoScaleπ: Scale-Aware Data Mixing for Pre-Training LLMs] Published as a conference paper at **COLM 2025*β¦β14Aug 8, 2025Updated last year
- Self-Supervised Alignment with Mutual Informationβ20May 24, 2024Updated 2 years ago
- Fast and Slow Generating: An Empirical Study on Large and Small Language Models Collaborative Decoding.β13Nov 19, 2024Updated last year
- Aioli: A unified optimization framework for language model data mixingβ34Jan 17, 2025Updated last year
- This respository is used for time reasoning task for mult-session dialogue system.β18Feb 7, 2026Updated 8 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- My Implementation of Q-Sparse: All Large Language Models can be Fully Sparsely-Activatedβ36Aug 14, 2024Updated 2 years ago
- A toolkit for automated alignment research.β16Jul 3, 2026Updated 3 months ago
- β20Nov 4, 2025Updated 11 months ago
- Official implementation of ECCV24 paper: POAβ24Aug 8, 2024Updated 2 years ago
- Implementation for the paper "Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite Learning"β11Jan 10, 2025Updated last year
- Official implementation of Privacy Implications of Retrieval-Based Language Models (EMNLP 2023). https://arxiv.org/abs/2305.14888β37Jun 10, 2024Updated 2 years ago
- "FiD-ICL: A Fusion-in-Decoder Approach for Efficient In-Context Learning" (ACL 2023)β15Jul 24, 2023Updated 3 years ago
- β25Dec 13, 2024Updated last year
- β13Jun 4, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- This is the oficial repository for "Safer-Instruct: Aligning Language Models with Automated Preference Data"β17Feb 22, 2024Updated 2 years ago
- β113Jul 15, 2025Updated last year
- ScienceMeter: Tracking Scientific Knowledge Updates in Language Models, COLM 2026β17Jun 28, 2025Updated last year
- Is In-Context Learning Sufficient for Instruction Following in LLMs? [ICLR 2025]β34Jan 23, 2025Updated last year
- DialOp: Decision-oriented dialogue environments for collaborative language agentsβ114Nov 15, 2024Updated last year
- CaMML:Context-Aware MultiModal Learner for Large Models (ACL 2024 SAC Award)β15May 21, 2025Updated last year
- GPTQ inference TVM kernelβ41Apr 25, 2024Updated 2 years ago
- We introduce EMMET and unify model editing with popular algorithms ROME and MEMIT.β28Dec 16, 2024Updated last year
- β47Jun 11, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The Code and Script of "David's Slingshot: A Strategic Coordination Framework of Small LLMs Matches Large LLMs in Data Synthesis"β34Jun 13, 2025Updated last year
- β56Jun 23, 2026Updated 3 months ago
- [AAAI 2024] History Matters: Temporal Knowledge Editing in Large Language Modelβ13Dec 17, 2023Updated 2 years ago
- β17Nov 20, 2024Updated last year
- β10Mar 13, 2023Updated 3 years ago
- Code for paper "Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System"β73Nov 14, 2024Updated last year
- Code for the arXiv preprint "The Unreasonable Effectiveness of Easy Training Data"β48Jan 17, 2024Updated 2 years ago
- [ACL 2024 (Oral)] A Prospector of Long-Dependency Data for Large Language Modelsβ61Jul 23, 2024Updated 2 years ago
- lanmt ebmβ12Jun 19, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Cramming 1568 Tokens into a Single Vector and Back Again: Exploring the Limits of Embedding Space Capacity (ACL 2025, oral)β35Jun 14, 2025Updated last year
- Decoding Attention is specially optimized for MHA, MQA, GQA and MLA using CUDA core for the decoding stage of LLM inference.β48Jun 11, 2025Updated last year
- A Data Source for Reasoning Embodied Agentsβ20Sep 18, 2023Updated 3 years ago
- Augmenting Statistical Models with Natural Language Parametersβ28Sep 17, 2024Updated 2 years ago
- Dateset Reset Policy Optimizationβ30Apr 12, 2024Updated 2 years ago
- Exploring Few-Shot Adaptation of Language Models with Tablesβ25Aug 22, 2022Updated 4 years ago
- Reproducible Language Agent Researchβ36Jun 25, 2025Updated last year