The official implementation of the paper "A Dual-Space Framework for General Knowledge Distillation of Large Language Models".
☆18Jan 4, 2026Updated 7 months ago
Alternatives and similar repositories for DSKDv2
Users that are interested in DSKDv2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repo for the EMNLP'24 Paper "Dual-Space Knowledge Distillation for Large Language Models". A general white-box KD framework for both same…☆64Mar 21, 2026Updated 5 months ago
- Code for "Think Natively: Unlocking Multilingual Reasoning with Consistency-Enhanced Reinforcement Learning".☆28Nov 11, 2025Updated 9 months ago
- ☆27Aug 31, 2025Updated 11 months ago
- [ICML 2026] Hybrid Policy Distillation (HPD) is a practical distillation framework for reasoning-oriented language models. This repositor…☆24Apr 24, 2026Updated 4 months ago
- A user-friendly & efficient knowledge distillation framework for LLMs, supporting off-policy, on-policy (OPD), cross-tokenizer, multimoda…☆246Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official implementation for "Knowledge Distillation with Refined Logits".☆24Aug 26, 2024Updated 2 years ago
- Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision☆18Apr 1, 2025Updated last year
- ☆23Aug 3, 2026Updated 3 weeks ago
- This is the official code for OThink-R1 project.☆21Jun 19, 2025Updated last year
- ☆17Jun 10, 2025Updated last year
- Python source code for EMNLP 2021 Findings paper: "Subword Mapping and Anchoring Across Languages".☆13Sep 17, 2021Updated 4 years ago
- ☆13Jul 2, 2025Updated last year
- An opinionated NLP research template☆10Aug 29, 2024Updated 2 years ago
- This repository contains code for the paper "Better Estimation of the KL Divergence Between Language Models"☆19May 30, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- The code of ACL2022 paper "Conditional Bilingual Mutual Information based Adaptive Training for Neural Machine Translation"..☆14Aug 6, 2022Updated 4 years ago
- ☆15Apr 11, 2024Updated 2 years ago
- Lowering PyTorch's Memory Consumption for Selective Differentiation☆12Aug 29, 2024Updated 2 years ago
- [NeurIPS 2024] Goldfish Loss: Mitigating Memorization in Generative LLMs☆98Nov 17, 2024Updated last year
- ☆16Jun 14, 2024Updated 2 years ago
- [ICANN 2024 (Oral)] MISS: A Generative Pre-training and Fine-tuning Approach for Med-VQA☆12Aug 8, 2024Updated 2 years ago
- ☆22Dec 11, 2024Updated last year
- ☆16Jul 23, 2024Updated 2 years ago
- IJCAI-18 阿里妈妈搜索广告转化预测大赛,top50方案☆15May 16, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Jan 14, 2026Updated 7 months ago
- EfficientRollout: System-Aware Self-Speculative Decoding for RL Rollouts☆16Jul 27, 2026Updated last month
- 2024CCF国际AIOps挑战赛-赛道二(GLM4):基于检索增强的运维知识问答挑战赛解决方案分享。☆14Jul 5, 2024Updated 2 years ago
- MCPToolBench++ MCP Model Context Protocol Tool Use Benchmark on AI Agent and Model Tool Use Ability☆44Mar 17, 2026Updated 5 months ago
- An Open-Source Knowledge-Enhanced Multilingual Supervised Fine-tuning Dataset☆27Jan 19, 2025Updated last year
- ☆15Jan 24, 2025Updated last year
- Source code of “Reinforcement Learning with Token-level Feedback for Controllable Text Generation (NAACL 2024)☆17Dec 8, 2024Updated last year
- Implementation and evaluation of Scaling Embedding Layers in Language Models research paper☆16Feb 2, 2026Updated 6 months ago
- This is the official leaderboard of the six practice for the new commers of BJTUNLPers.☆15Dec 17, 2019Updated 6 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Entropy-Driven GRPO with Guided Error Correction for Advantage Diversity☆22Aug 28, 2025Updated last year
- ☆22Dec 15, 2023Updated 2 years ago
- [ACL 2025 Main] Official Pytorch Implementation for "State-offset Tuning: State-based Parameter-Efficient Fine-Tuning for State Space Mod…☆15Jun 9, 2025Updated last year
- Official implementation of ECCV24 paper: POA☆24Aug 8, 2024Updated 2 years ago
- Official Code for "Learning to Reason via Mixture-of-Thought for Logical Reasoning"☆29Nov 20, 2025Updated 9 months ago
- Beyond Myopia: Learning from Positive and Unlabeled Data through Holistic Predictive Trends [NeurIPS 2023]☆10Jan 28, 2024Updated 2 years ago
- The official implementation for "Mitigating Overthinking in Large Reasoning Models via Manifold Steering"☆15May 29, 2025Updated last year