[NeurIPS 2026] AsyncOPD: How Stale Can On-Policy Distillation Be?
☆36Jun 29, 2026Updated 3 months ago
Alternatives and similar repositories for async-opd
Users that are interested in async-opd are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2026] EfficientRollout: System-Aware Self-Speculative Decoding for RL Rollouts☆16Updated this week
- [ACL 2025 Main] Official Pytorch Implementation for "State-offset Tuning: State-based Parameter-Efficient Fine-Tuning for State Space Mod…☆15Jun 9, 2025Updated last year
- ☆20May 26, 2026Updated 4 months ago
- [WACV 2025 Oral] Counting Guidance for High Fidelity Text-to-Image Synthesis☆15Jul 4, 2025Updated last year
- [ICML 2025] Parameter-Efficient Fine-Tuning of State Space Models☆25Jun 9, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR 2026] Draft-based Approximate Inference for LLMs☆21Mar 10, 2026Updated 6 months ago
- [ICLR 2026] ParallelBench: Understanding the Tradeoffs of Parallel Decoding in Diffusion LLMs☆48Mar 27, 2026Updated 6 months ago
- UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation☆16Aug 12, 2025Updated last year
- CDLM: Consistency Diffusion Language Models for Faster Sampling☆41Nov 25, 2025Updated 10 months ago
- Implementation and dataset for paper "Can MLLMs Perform Text-to-Image In-Context Learning?"☆48Jun 2, 2025Updated last year
- [ECCV 2024] Official Pytorch Implementation for "Eta Inversion: Designing an Optimal Eta Function for Diffusion-based Real Image Editing"☆34Jun 16, 2025Updated last year
- Code for "The Expressive Power of Low-Rank Adaptation".☆20Apr 19, 2024Updated 2 years ago
- [ICML 2026] Hybrid Policy Distillation (HPD) is a practical distillation framework for reasoning-oriented language models. This repositor…☆24Apr 24, 2026Updated 5 months ago
- 자연어 처리 기반 [한글 서술형 수학문제 데이터셋] 공개 저장소입니다.☆14Jun 12, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- repository for Causal&NLP reading group☆11Jan 30, 2026Updated 8 months ago
- Official codebase for the paper "WorldMemArena: Evaluating Multimodal Agent Memory Through Action–World Interaction"☆28May 29, 2026Updated 4 months ago
- PANDA: Prompt Transfer Meets Knowledge Distillation for Efficient Model Adaptation☆16Mar 28, 2023Updated 3 years ago
- Software for creating jigsaw puzzles using LaTeX, with output similar to Tarsia's Formulator software☆14Feb 8, 2021Updated 5 years ago
- Official Implementation of Trajectory-Refined Distillation☆37Jun 9, 2026Updated 3 months ago
- ☆13Dec 7, 2022Updated 3 years ago
- Simple Light OS source repository☆22Aug 10, 2025Updated last year
- [ICLR'24 spotlight] Tool-Augmented Reward Modeling☆53Jun 6, 2025Updated last year
- Code for Language-Interfaced FineTuning for Non-Language Machine Learning Tasks.☆135Nov 11, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Efficient and Effective Weight-Ensembling Mixture of Experts for Multi-Task Model Merging. Arxiv, 2024.☆16Oct 28, 2024Updated last year
- Source Code for our ICLR'26 paper☆18Feb 22, 2026Updated 7 months ago
- [ICML 2026] Spherical Steering: Geometry-Aware Activation Rotation for Language Models☆23May 19, 2026Updated 4 months ago
- code for ACL2024-main: BatchEval: Towards Human-like Text Evaluation☆19May 20, 2024Updated 2 years ago
- Fine-Grained Causal Dynamics Learning with Quantization for Improving Robustness in Reinforcement Learning (ICML 2024)☆20Jun 5, 2024Updated 2 years ago
- ☆15May 4, 2024Updated 2 years ago
- Structural Causal Bandit☆28Sep 6, 2026Updated 3 weeks ago
- Official GitHub repo for Scaling Physical Reasoning with the PHYSICS Dataset (NeurIPS25).☆16Sep 20, 2025Updated last year
- a common lisp library for data analysis and manipulation☆12May 28, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Framework to generate observational and interventional samples from structural equation models (SEMs)☆22Apr 14, 2025Updated last year
- Common Lisp implementation of the Erlang External Term Format☆16Dec 31, 2022Updated 3 years ago
- [ICLR 2026] Official Implementation of ProxyThinker: Test-Time Guidance through Small Visual Reasoners.☆22Sep 24, 2025Updated last year
- Multi-Objective Causal Bayesian Optimisation, a new paradigm for finding Pareto-optimal interventions in multi-outcome causal models☆18Jun 2, 2025Updated last year
- The implementation for FREE-Merging: Fourier Transform for Model Merging with Lightweight Experts (ICCV25)☆16Jun 26, 2025Updated last year
- [arXiv] "Linear Dynamics in the RLVR Training of Large Language Models"☆20May 25, 2026Updated 4 months ago
- 자체 구축한 한국어 평가 데이터셋을 이용한 한국어 모델 평가☆31May 31, 2024Updated 2 years ago