EfficientRollout: System-Aware Self-Speculative Decoding for RL Rollouts
☆16Jun 24, 2026Updated 3 weeks ago
Alternatives and similar repositories for EfficientRollout
Users that are interested in EfficientRollout are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆20May 26, 2026Updated last month
- ☆15Aug 11, 2025Updated 11 months ago
- [ICLR 2026] ParallelBench: Understanding the Tradeoffs of Parallel Decoding in Diffusion LLMs☆46Mar 27, 2026Updated 3 months ago
- CDLM: Consistency Diffusion Language Models for Faster Sampling☆42Nov 25, 2025Updated 7 months ago
- UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation☆17Aug 12, 2025Updated 11 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Anatomy of High-Performance GEMM with Online Fault Tolerance on GPUs☆14Apr 3, 2025Updated last year
- ☆11Oct 3, 2022Updated 3 years ago
- This repository contains papers for a comprehensive survey on accelerated generation techniques in Large Language Models (LLMs).☆11May 24, 2024Updated 2 years ago
- Learning from Mixed Rollouts: Logit Fusion as a Bridge Between Imitation and Exploration☆17Feb 24, 2026Updated 4 months ago
- Multimodal extreme classification☆21May 1, 2024Updated 2 years ago
- ☆18Jan 17, 2024Updated 2 years ago
- The official implementation of the paper "A Dual-Space Framework for General Knowledge Distillation of Large Language Models".☆18Jan 4, 2026Updated 6 months ago
- Explainable Interactive Concept Learning☆15Mar 26, 2023Updated 3 years ago
- Tensorflow implementation of the `intelligent synapse' model from [Zenke et al., (2017)] and application to the Permuted MNIST benchmark.☆22Aug 2, 2017Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A Concept-Centric Framework for Intelligent Agents☆27Oct 1, 2025Updated 9 months ago
- [NeurIPS'21] RoMA: Robust Model Adaptation for Offline Model-based Optimization☆15Oct 28, 2021Updated 4 years ago
- Updated figures for "A benchmarking of WGS-based structural variant callers" paper☆27Apr 3, 2022Updated 4 years ago
- [NeurIPS 2025] Multipole Attention for Efficient Long Context Reasoning☆24Dec 5, 2025Updated 7 months ago
- Prebuilt binaries for Code-OSS on Win32-arm64. Essential build instructions also included.☆23Apr 11, 2020Updated 6 years ago
- Discretized Integrated Gradients for Explaining Language Models (EMNLP 2021)☆27Mar 26, 2022Updated 4 years ago
- Multimodal Graph Network (MGN): Code repo, examples from the paper☆25Apr 30, 2021Updated 5 years ago
- Some microbenchmarks and design docs before commencement☆11Feb 1, 2021Updated 5 years ago
- A sample app to debug and validate cellular modems on balena devices☆13Jun 5, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for our ACL '23 paper titled "Grokking of Hierarchical Structure in Vanilla Transformers"☆26Oct 8, 2023Updated 2 years ago
- ☆16Jul 29, 2025Updated 11 months ago
- Code for 'Why is Winoground Hard? Investigating Failures in Visuolinguistic Compositionality', EMNLP 2022☆31May 29, 2023Updated 3 years ago
- ☆15Jun 28, 2023Updated 3 years ago
- [Poster; ICLR 2026] [Oral; Neurips OPT2024] μLO: Compute-Efficient Meta-Generalization of Learned Optimizers☆16Apr 15, 2026Updated 3 months ago
- Repository for the code and dataset for the paper: "Have LLMs Advanced enough? Towards Harder Problem Solving Benchmarks For Large Langu…☆39Dec 18, 2023Updated 2 years ago
- A simple one file python script that executes AI processes defined in YML.☆14Mar 26, 2023Updated 3 years ago
- [CVPR23 Highlight] CREPE: Can Vision-Language Foundation Models Reason Compositionally?☆35Apr 27, 2023Updated 3 years ago
- Let's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning with LLMs☆41Jan 30, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 🌿快速生成文件夹目录结构,支持定义目录层级,支持生成到 markdown 文件。☆13Oct 19, 2022Updated 3 years ago
- ☆34Apr 19, 2024Updated 2 years ago
- ☆17Apr 7, 2025Updated last year
- Tutorial: Writing R and Python Packages with Multithreaded C++ Code using BLAS, AVX2/AVX512, OpenMP, C++11 Threads and Cuda GPU accelerat…☆13Nov 27, 2022Updated 3 years ago
- A list of papers and other resources on language-guided image editing.☆39Jan 13, 2021Updated 5 years ago
- Are Intermediate Layers and Labels Really Necessary? A General Language Model Distillation Method ; GKD: A General Knowledge Distillation…☆34Aug 4, 2023Updated 2 years ago
- Example Express.js app connecting to PlanetScale☆29May 6, 2026Updated 2 months ago