☆23Oct 10, 2025Updated 10 months ago
Alternatives and similar repositories for Small-Model-Learnability-Gap
Users that are interested in Small-Model-Learnability-Gap are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Jul 31, 2025Updated last year
- [ACL 2026 Main] Official Repo for Paper "Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Ali…☆17Jul 1, 2026Updated last month
- ☆55May 22, 2025Updated last year
- Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts☆26Feb 23, 2024Updated 2 years ago
- [ACL 2025 Findings] Official implementation of the paper "Unveiling the Key Factors for Distilling Chain-of-Thought Reasoning".☆23Feb 26, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 此项目是我个人对MIT 6.5940 课程作业的答案,学习笔记和心得。☆15Mar 1, 2024Updated 2 years ago
- ☆15May 27, 2025Updated last year
- Official implement of CIKM2025: 《UniECS: Unified Multimodal E-Commerce Search Framework with Gated Cross-modal Fusion》☆21Sep 17, 2025Updated 10 months ago
- ☆20Jun 17, 2024Updated 2 years ago
- EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems (ICLR'26)☆28Nov 3, 2025Updated 9 months ago
- Kaggle AIMO2 solution with token-efficient reasoning LLM recipes☆51Aug 7, 2025Updated last year
- R1-Code-Interpreter: Training LLMs to Reason with Code via Supervised and Reinforcement Learning☆45Feb 9, 2026Updated 6 months ago
- Analysing fairness of graph anomaly detection methods☆12Nov 24, 2024Updated last year
- Learning MLPs to replace GNN☆10Jun 3, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A curated list of cutting-edge research papers and resources on Long Chain-of-Thought (CoT) Reasoning with Tools.☆46Dec 17, 2025Updated 7 months ago
- A Chinese Character BERT Trained with Multi-Level Masking☆13Sep 24, 2023Updated 2 years ago
- ☆11May 17, 2024Updated 2 years ago
- A toolkit for automated alignment research.☆15Jul 3, 2026Updated last month
- ☆46Mar 4, 2025Updated last year
- A Workbench for Autograding Retrieve/Generate Systems☆15Jun 30, 2025Updated last year
- ☆43May 6, 2024Updated 2 years ago
- Jointly Optimizing Large Language Models for Reasoning and Self-Refinement☆15Apr 22, 2026Updated 3 months ago
- Codebase for character-centric story understanding☆14Jan 20, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- personal settings for linux tools, including zsh, vim, tmux, pip.☆11Dec 2, 2019Updated 6 years ago
- ☆73Oct 23, 2025Updated 9 months ago
- Source code for the paper 'Uncovering Neural Scaling Laws in Molecular Representation Learning' (NeurIPS 2023 Datasets and Benchmarks).☆14Dec 2, 2023Updated 2 years ago
- ☆13Feb 8, 2025Updated last year
- ☆11May 1, 2022Updated 4 years ago
- SMART introduces a novel test-time framework where Small Language Models (SLMs) reason step-by-step, and Large Language Models (LLMs) pro…☆12Jul 9, 2025Updated last year
- Code for the paper Xiangqi-R1: Enhancing Spatial Strategic Reasoning in LLMs for Chinese Chess via Reinforcement Learning☆15Jul 23, 2025Updated last year
- ☆19Jul 30, 2025Updated last year
- Official resources of "The First Few Tokens Are All You Need: An Efficient and Effective Unsupervised Prefix Fine-Tuning Method for Reaso…☆20Jun 13, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Documentation at☆14Mar 27, 2025Updated last year
- Awesome-Parallel-Reasoning: Unlocking the reasoning potential of LLMs. Papers, Code, Resources & Survey.☆55Mar 8, 2026Updated 5 months ago
- SophiaVL-R1: Reinforcing MLLMs Reasoning with Thinking Reward☆94Aug 8, 2025Updated last year
- Source codes for the paper "Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning" (PDMER) which p…☆14Mar 24, 2025Updated last year
- 欢迎参加中文讽刺计算评测任务!☆14Nov 4, 2024Updated last year
- ☆46Dec 30, 2024Updated last year
- Model for multi-turn tool calling of bash functions☆22Jan 26, 2026Updated 6 months ago