OPRD: On-Policy Representation Distillation (https://arxiv.org/abs/2606.06021)
☆73Sep 9, 2026Updated this week
Alternatives and similar repositories for OPRD
Users that are interested in OPRD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe☆1,003Aug 20, 2026Updated 3 weeks ago
- [NeurIPS 2025] Universal Few-Shot Spatial Control for Diffusion Models☆21Sep 18, 2025Updated 11 months ago
- Official repository for the paper "Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation"☆294May 28, 2026Updated 3 months ago
- ☆18Apr 6, 2026Updated 5 months ago
- A curated collection of papers and resources on On-Policy Distillation for Large Language Models.☆540Aug 12, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ABench is an evolving open-source benchmark suite designed to rigorously evaluate and enhance Large Language Models (LLMs) on complex cro…☆29Jul 30, 2026Updated last month
- A new quantization framwork☆16Nov 19, 2025Updated 9 months ago
- [ICLR 2026 Oral] RAIN-Merging☆15Mar 9, 2026Updated 6 months ago
- Implementation of Gradient Information Optimization (GIO) for effective and scalable training data selection☆14Jun 22, 2023Updated 3 years ago
- Awesome List for On-Policy Distillation☆857Aug 26, 2026Updated 2 weeks ago
- [EMNLP 2024] Quantize LLM to extremely low-bit, and finetune the quantized LLMs☆17Jul 18, 2024Updated 2 years ago
- [ICCV 2025 Highlight] Less is More: Empowering GUI Agent with Context-Aware Simplification☆48Mar 12, 2026Updated 6 months ago
- PyTorch code for the CVPR'23 paper: "ConStruct-VL: Data-Free Continual Structured VL Concepts Learning"☆14Feb 5, 2024Updated 2 years ago
- Official PyTorch implementation of RACRO (https://www.arxiv.org/abs/2506.04559)☆19Jul 1, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- On Policy Distillation Build on top of Verl☆98Sep 3, 2026Updated last week
- Code for MERL's ECCV 2022 paper on Cross-Modal Knowledge Transfer Without Task-Relevant Source Data☆11Jul 19, 2022Updated 4 years ago
- ☆12Jan 30, 2024Updated 2 years ago
- LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric Rewards☆40Jun 1, 2026Updated 3 months ago
- This repo lists some researches and applications in PU learning.☆12Mar 12, 2020Updated 6 years ago
- ☆35Apr 22, 2026Updated 4 months ago
- [EMNLP2026 Main] Official implementation of “Domino: Decoupling Causal Modeling from Autoregressive Drafting in Speculative Decoding”.☆138Jul 25, 2026Updated last month
- ☆13Oct 24, 2023Updated 2 years ago
- Codebase for ICML submission "DOGE: Domain Reweighting with Generalization Estimation"☆21Feb 29, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ACL-26 (main)] From Verbatim to Gist Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video A…☆41Apr 19, 2026Updated 4 months ago
- ☆82May 12, 2026Updated 4 months ago
- [CVPR 24] This is official implication for our paper: ''CroSel: Cross Selection of Confident Pseudo Labels for Partial-Label Learning''.☆15Apr 27, 2025Updated last year
- ☆15May 4, 2024Updated 2 years ago
- Signed distance shader library for Unity (including shader graph)☆12Nov 4, 2025Updated 10 months ago
- [ECCV 2024] Reliable Spatial-Temporal Voxels for Multi-Modal Test-Time Adaptation☆19Jan 12, 2026Updated 8 months ago
- The code for the paper "Efficient Self-Supervised Video Hashing with Selective State Spaces" (AAAI'25).☆24Aug 2, 2025Updated last year
- A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models☆813Aug 30, 2026Updated 2 weeks ago
- Operation System Course's Educoder excrises shell script. / 操作系统课程的头歌过关脚本。☆11Jun 16, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆16Jun 4, 2024Updated 2 years ago
- ☆19Jun 26, 2024Updated 2 years ago
- ☆20Oct 13, 2024Updated last year
- [ICML 2024] SPP: Sparsity-Preserved Parameter-Efficient Fine-Tuning for Large Language Models☆22May 28, 2024Updated 2 years ago
- Official code for ICLR 2024 paper, SEABO: A Simple Search-Based Method for Offline Imitation Learning☆13Jan 19, 2024Updated 2 years ago
- Official Code for IJCAI25 paper "Zero-shot Generalist Graph Anomaly Detection with Unified Neighborhood Prompts"☆23Jun 7, 2025Updated last year
- This repository contains the code for the paper “Neuro-Symbolic Query Compiler”, accepted to the Findings of ACL 2025.☆19Oct 20, 2025Updated 10 months ago