☆23Oct 10, 2025Updated 9 months ago
Alternatives and similar repositories for Small-Model-Learnability-Gap
Users that are interested in Small-Model-Learnability-Gap are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Jul 31, 2025Updated 11 months ago
- [ACL 2026 Main] Official Repo for Paper "Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Ali…☆17Jul 1, 2026Updated 3 weeks ago
- ☆16Dec 3, 2024Updated last year
- [ACL '26] source code for the paper: "Long-Chain Reasoning Distillation via Adaptive Prefix Alignment"☆16Jan 21, 2026Updated 6 months ago
- Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts☆26Feb 23, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ACL 2025 Findings] Official implementation of the paper "Unveiling the Key Factors for Distilling Chain-of-Thought Reasoning".☆22Feb 26, 2025Updated last year
- 此项目是我个人对MIT 6.5940 课程作业的答案,学习笔记和心得。☆15Mar 1, 2024Updated 2 years ago
- Official implement of CIKM2025: 《UniECS: Unified Multimodal E-Commerce Search Framework with Gated Cross-modal Fusion》☆21Sep 17, 2025Updated 10 months ago
- ☆15May 27, 2025Updated last year
- COS-PLAY: Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Game Play☆29Jul 11, 2026Updated last week
- EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems (ICLR'26)☆24Nov 3, 2025Updated 8 months ago
- ☆20Jun 17, 2024Updated 2 years ago
- Reasoning and Tool-use Compete in Agentic RL: From Quantifying Interference to Disentangled Tuning