A curated list of resources on on-policy distillation
☆25Apr 13, 2026Updated 5 months ago
Alternatives and similar repositories for awesome-on-policy-distillation
Users that are interested in awesome-on-policy-distillation are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2025] The Indra Representation Hypothesis for Multimodal Alignment☆32Feb 3, 2026Updated 7 months ago
- [EMNLP 2025] Representation Potentials of Foundation Models for Multimodal Alignment: A Survey☆33Feb 3, 2026Updated 7 months ago
- [NeurIPS 2023] Latent Graph Inference with Limited Supervision☆33Feb 1, 2024Updated 2 years ago
- [ACL 2026] Paper list of Video LLM hallucination. Welcome to Star and Contribute!☆42Sep 1, 2026Updated last week
- SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting☆29Jun 22, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Generalizing from SIMPLE to HARD Visual Reasoning: Can We Mitigate Modality Imbalance in VLMs?☆19Jun 3, 2025Updated last year
- [ICLR 2026🔥] SHIELD: Suppressing Hallucinations In LVLM Encoders via Bias and Vulnerability Defense☆19Mar 24, 2026Updated 5 months ago
- Open-source RL Framework with Online Teacher-Student Distillation☆22Mar 5, 2026Updated 6 months ago
- ThinkJEPA: Empowering Latent World Models with Large Vision-Language Reasoning Model☆54Jul 19, 2026Updated last month
- Awesome List for On-Policy Distillation☆857Aug 26, 2026Updated 2 weeks ago
- Policy Optimization is awesome, let’s put a tree on it! 🌲🌟☆22Jul 4, 2025Updated last year
- [ACL 2026] Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning☆94Jan 22, 2026Updated 7 months ago
- NetEaseCrowd dataset, a collection of data obtained from You Ling crowdsourcing platform, Fuxi AI Lab, NetEase.☆14Dec 19, 2024Updated last year
- A survey on MM-LLMs for long video understanding: From Seconds to Hours: Reviewing MultiModal Large Language Models on Comprehensive Long…☆25Sep 12, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for ProTrix: Building Models for Planning and Reasoning over Tables with Sentence Context☆17Nov 15, 2024Updated last year
- ZJU毛概资料汇总☆14Mar 16, 2024Updated 2 years ago
- Disentangled Representation Learning for Recommendation, TPAMI 2022☆11Oct 11, 2024Updated last year
- ☆631Sep 5, 2026Updated last week
- Persistent Topological Features in Large Language Models☆19Jul 23, 2026Updated last month
- RL Recommendation System☆13Aug 30, 2019Updated 7 years ago
- ☆13Oct 5, 2022Updated 3 years ago
- Attention-based multimodal fusion for sentiment analysis☆13Aug 14, 2018Updated 8 years ago
- CRNN with Self-Attention☆10Apr 8, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The simulator for education☆17Mar 16, 2021Updated 5 years ago
- 基于苏剑林项目的复用,应用于金融事件关系抽取☆11Mar 26, 2021Updated 5 years ago
- ☆67Jul 3, 2026Updated 2 months ago
- ☆17Mar 17, 2026Updated 5 months ago
- Pytorch plugin to generate saliency maps for neural networks☆12Nov 1, 2018Updated 7 years ago
- 🎯 Collect and create an outstanding homepage template that is relevant to your project. 🎀 Create a homepage for your work!☆18Sep 3, 2026Updated last week
- ☆10Apr 9, 2021Updated 5 years ago
- szuthesis 深圳大学学位论文 LaTeX 模板☆34Feb 23, 2021Updated 5 years ago
- Mixture of Lora Experts☆11Apr 7, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [AAAI 2026] CrossVid: A Comprehensive Benchmark for Evaluating Cross-Video Reasoning in Multimodal Large Language Models☆23Jul 9, 2026Updated 2 months ago
- ☆14Jul 17, 2025Updated last year
- TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models☆17Jan 2, 2025Updated last year
- A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models☆813Aug 30, 2026Updated 2 weeks ago
- ZJU-2019-Computer Network☆13Aug 3, 2020Updated 6 years ago
- [CVPR 2026] Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens☆298Aug 2, 2025Updated last year
- [NeurIPS 2025] Neural Discrete Token Representation Learning for Extreme Token Reduction in Video Large Language Models☆17Aug 19, 2026Updated 3 weeks ago