☆18Dec 22, 2025Updated 8 months ago
Alternatives and similar repositories for verl
Users that are interested in verl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of On-Policy Distillation (GKD) for Language Models - ICLR 2024☆24Nov 24, 2025Updated 9 months ago
- MCOUT: Multimodal Chain of Continuous Thought for Latent Reasoning☆22Oct 4, 2025Updated 10 months ago
- [ICLR 2025] Official repository for the paper "Influence-Guided Diffusion for Dataset Distillation".☆15Feb 12, 2025Updated last year
- Code and dataset for paper "SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition"☆19Mar 17, 2026Updated 5 months ago
- This repo contains demo ROS code based on Control-Toolbox and ACADO Toolkit☆14Feb 12, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ACL21 Math Word Problem Solving with Explicit Numerical Values☆13Nov 10, 2021Updated 4 years ago
- WorldReasonBench: Human-Aligned Stress Testing of Video Generators as Future World-State Predictors☆23May 19, 2026Updated 3 months ago
- This repository provides the code for applying Contrastive Learning Penalty Loss (CLPL) and Mixture of Experts (MoE) to the BGE-M3 text e…☆11Dec 27, 2024Updated last year
- This is tensorflow 2.2 based SCAMET framework for remote sensing image captioning.☆13Aug 10, 2023Updated 3 years ago
- This is a fork of SGLang for hip-attention integration. Please refer to hip-attention for detail.☆18Mar 31, 2026Updated 4 months ago
- Implementation of some active contour model / Snake algorithms☆14Jan 4, 2018Updated 8 years ago
- pure go for rwkv☆18Dec 31, 2023Updated 2 years ago
- ☆16Jul 29, 2025Updated last year
- Official code for paper "Revisiting Model Interpolation for Efficient Reasoning"☆17Jul 14, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- WDEL是一个基于Wikidata知识库的实体链接系统。☆11Feb 12, 2025Updated last year
- PHP with FPM Dockerfile for trusted automated Docker builds.☆12Mar 2, 2016Updated 10 years ago
- ☆16Apr 30, 2025Updated last year
- LoongRL: Reinforcement Learning for Advanced Reasoning over Long Contexts (ICLR 2026 Oral)☆36Feb 20, 2026Updated 6 months ago
- This repo contains the source code for reproducing the experimental results in semantic density paper (Neurips 2024)☆21Sep 28, 2025Updated 11 months ago
- ☆21Jan 25, 2021Updated 5 years ago
- Overflow Prevention Enhances Long-Context Recurrent LLMs (COLM 2025)☆18Jul 8, 2025Updated last year
- ☆21Oct 12, 2021Updated 4 years ago
- PyTorch tool for training with bigger batch size on the GPU☆11Feb 26, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆17Jan 31, 2025Updated last year
- BERT&RoBERTa预训练代码,tensorflow和torch两种版本实现☆13Feb 8, 2023Updated 3 years ago
- handy tools for user study☆21May 21, 2024Updated 2 years ago
- FastAPI to serve Qwen-ASR with streaming support. Tested. Benchmarked. Flash Attention 2. Fast & Stable.☆15Jun 24, 2026Updated 2 months ago
- The official implemention of "Depth-Breadth Synergy in RLVR: Unlocking LLM Reasoning Gains with Adaptive Exploration" (ICML 2026)☆24Feb 4, 2026Updated 6 months ago
- ☆14Mar 11, 2025Updated last year
- pytorch版simcse无监督语义相似模型☆22May 13, 2021Updated 5 years ago
- ☆21Apr 16, 2025Updated last year
- Low-Rank Llama Custom Training☆23Mar 27, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official repository of paper "Parameters vs. Context: Fine-Grained Control of Knowledge Reliance in Language Models"☆26May 27, 2025Updated last year
- Official repository for the paper "Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation"☆291May 28, 2026Updated 3 months ago
- Python library to compress LitGPT models for resource efficient inference.☆16Jul 31, 2026Updated 3 weeks ago
- An Ultra-Long Output Reinforcement Learning Approach☆23Jul 31, 2025Updated last year
- ☆27Jun 10, 2025Updated last year
- ☆14Apr 23, 2020Updated 6 years ago
- ☆25Oct 10, 2023Updated 2 years ago