Dogfooding vLLM backport for older GPUs.
☆295Oct 6, 2026Updated this week
Alternatives and similar repositories for vllm-backport
Users that are interested in vllm-backport are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS'25] Beyond Accuracy: Dissecting Mathematical Reasoning for LLMs Under Reinforcement Learning☆17Dec 12, 2025Updated 9 months ago
- ☆12Oct 10, 2024Updated 2 years ago
- Use contrastive learning to train a large language model (LLM) as a retriever☆12Jul 19, 2024Updated 2 years ago
- [ICML 2024] Junk DNA Hypothesis: A Task-Centric Angle of LLM Pre-trained Weights through Sparsity; Lu Yin*, Ajay Jaiswal*, Shiwei Liu, So…☆16Apr 21, 2025Updated last year
- Official implementation of "RDD: Retrieval-Based Demonstration Decomposer for Planner Alignment in Long-Horizon Tasks" NeurIPS 2025.☆16Jul 22, 2026Updated 2 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆22Nov 17, 2025Updated 10 months ago
- ☆17Aug 6, 2024Updated 2 years ago
- This repository collects various works that reproduce DeepSeek R1, as well as works related to DeepSeek R1 and the DeepSeek series.☆20Apr 27, 2025Updated last year
- [ICCV 2025] FVP: 4D Visual Pre-training for Robot Learning☆17Sep 5, 2025Updated last year
- Experiments with reasoning models, training techniques, papers☆30Updated this week
- ☆25Apr 8, 2026Updated 6 months ago
- ☆25Apr 25, 2025Updated last year
- BenchPush is a comprehensive benchmarking suite designed for mobile robots performing pushing-based tasks. It provides simulated environm…☆22Feb 12, 2026Updated 7 months ago
- [COLM 2025] EvalTree: Profiling Language Model Weaknesses via Hierarchical Capability Trees☆31Jul 11, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- LongProc: Benchmarking Long-Context Language Models on Long Procedural Generation☆36Feb 26, 2026Updated 7 months ago
- ☆32Sep 20, 2026Updated 3 weeks ago
- Reinforcement learning framework for training N2 humanoid robots in Isaac Gym. Includes environment definitions, motion loaders, AMP pipe…☆25Nov 21, 2025Updated 10 months ago
- [NeurIPS 2023 Spotlight] Temperature Balancing, Layer-wise Weight Analysis, and Neural Network Training☆37Apr 7, 2025Updated last year
- A collection on the recent reproduction papers and projects on DeepSeek-R1☆31Feb 27, 2025Updated last year
- MemoChat: Tuning LLMs to Use Memos for Consistent Long-Range Open-Domain Conversation☆30Apr 18, 2024Updated 2 years ago
- A collection of long-term memory papers☆40Jan 18, 2026Updated 8 months ago
- Code repository for "RL Grokking Recipe: How RL Unlocks and Transfers New Algorithms in LLMs""☆35Oct 12, 2025Updated 11 months ago
- FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions☆57Jul 3, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official repository for "Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress," presented at Co…☆36Feb 6, 2025Updated last year
- [CoRL 2025] Robot Learning from Any Images☆35Nov 11, 2025Updated 11 months ago
- [ICRA 2025] RACER: Rich Language-Guided Failure Recovery Policies for Imitation Learning☆46Oct 10, 2024Updated 2 years ago
- MemoryEQA☆28May 4, 2026Updated 5 months ago
- A Comprehensive Benchmarking Framework for Long-Term Conversational Memory Layers☆49Sep 3, 2026Updated last month
- ☆53May 11, 2025Updated last year
- Code record for MIT 6.S081 xv6-riscv lab 2021, some differences refer to the repository "xv6-lab-2020" https://github.com/NebulorDang/xv6…☆23Dec 2, 2021Updated 4 years ago
- [ACL 2026] From 1,000,000 Users to Every User: Scaling Up Personalized Preference for User-level Alignment☆40Jan 8, 2026Updated 9 months ago
- [ICLR 2025] BRIGHT: A Realistic and Challenging Benchmark for Reasoning-Intensive Retrieval☆215Sep 13, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [NeurIPS 2025] Official code repository for "Failure Prediction at Runtime for Generative Robot Policies".☆56Nov 3, 2025Updated 11 months ago
- RoboFAC: A Comprehensive Framework for Robotic Failure Analysis and Correction☆46Jul 15, 2026Updated 2 months ago
- General Reasoner: Advancing LLM Reasoning Across All Domains [NeurIPS25]☆232Nov 27, 2025Updated 10 months ago
- Watch Every Step! LLM Agent Learning via Iterative Step-level Process Refinement (EMNLP 2024 Main Conference)☆67Oct 18, 2024Updated last year
- Official implementation of FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment☆58Mar 24, 2026Updated 6 months ago
- [CVPR 2025] VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation☆53Apr 10, 2026Updated 6 months ago
- Official repository for ACL 2025 paper "Model Extrapolation Expedites Alignment"☆75May 20, 2025Updated last year