☆18Feb 24, 2026Updated 6 months ago
Alternatives and similar repositories for vllm-learn
Users that are interested in vllm-learn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A fault-tolerant RDMA-based disaggregated key-value store with 1-RTT UPDATEs and GETs thanks to the SWARM replication protocol☆14Sep 25, 2024Updated last year
- ☆16Feb 8, 2024Updated 2 years ago
- ☆16Apr 13, 2024Updated 2 years ago
- Mirror of the Xen Repository (PRs not accepted see: http://wiki.xenproject.org/wiki/Submitting_Xen_Project_Patches)☆17Sep 12, 2017Updated 8 years ago
- Official code and resources for the paper "EXIT: Context-Aware Extractive Compression for Enhancing Retrieval-Augmented Generation."☆26Jul 15, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- springboot集成各个组件,mybatis、rabbitmq、redis、elk、ldap、mqtt、websocket、socketio等,spring5新特性响应式webflux☆14Nov 8, 2023Updated 2 years ago
- ☆12Mar 13, 2023Updated 3 years ago
- some hexagon intrinsic examples based on Qualcomm Hexagon☆18Mar 7, 2025Updated last year
- A pure userspace filesystem developed on SPDK's block device layer, optimized for high-speed storage devices (NVMe/PMEM) in high-concurre…☆26Mar 11, 2026Updated 5 months ago
- 领域驱动设计——实战落地代码,基于单体项目拆分【基于SpringBootCloud的微服务项目】☆13Sep 21, 2022Updated 3 years ago
- 做Web3世界的巴别塔|看完你就懂什么是Web3|面向萌新的Web3“白皮书” |@Web3-Club☆16Aug 29, 2024Updated 2 years ago
- Windows paravirtualized☆27Sep 5, 2025Updated 11 months ago
- ☆11Oct 31, 2024Updated last year
- java面试相关内容☆15Nov 2, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Drawing Diagram tool. Which is made by angular 6.0.8 and mxGraph☆14Jul 25, 2018Updated 8 years ago
- ☆11Jun 6, 2023Updated 3 years ago
- 收录我原创文章的相关代码☆14Jul 1, 2022Updated 4 years ago
- A C++ port of karpathy/micrograd, a tiny scalar-valued autograd engine and a neural net library☆13Nov 24, 2023Updated 2 years ago
- Artifact evaluation repo for EuroSys'24.☆30Nov 7, 2023Updated 2 years ago
- A c++ hash map/table which utilizes simd (specifically Intel x86 SSE/AVX)☆12Apr 30, 2019Updated 7 years ago
- ☆27May 30, 2025Updated last year
- A C++-based RPC framework☆12Oct 28, 2021Updated 4 years ago
- A maze generator and path finder in Angular☆14Feb 19, 2026Updated 6 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 可运行的Claude代码源码,采用双端模型进行驱动☆18Apr 5, 2026Updated 4 months ago
- Large-Scale Disk-Based Vector Index☆45Jun 18, 2026Updated 2 months ago
- ☆17Jan 3, 2025Updated last year
- ☆11Apr 25, 2021Updated 5 years ago
- OneFlow Serving☆20Apr 10, 2025Updated last year
- ☆33Sep 29, 2021Updated 4 years ago
- Multiple GEMM operators are constructed with cutlass to support LLM inference.☆20Aug 3, 2025Updated last year
- LazyLog: A New Shared Log Abstraction for Low-Latency Applications☆47Apr 28, 2025Updated last year
- xiaozhi(小智AI)fluent风格桌面端应用☆27Oct 10, 2025Updated 10 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A torch compile backend for multi-targets☆51May 27, 2026Updated 3 months ago
- An IR for efficiently simulating distributed ML computation.☆33Jan 13, 2024Updated 2 years ago
- transformer tokenizers (e.g. BERT tokenizer) in C++ (WIP)☆18Apr 7, 2022Updated 4 years ago
- Pytorch implementation of our paper accepted by ICML 2024 -- CaM: Cache Merging for Memory-efficient LLMs Inference☆51Jun 19, 2024Updated 2 years ago
- Semantic-Aware Fine-Grained Correspondence, at ECCV 2022 (Oral)☆14Oct 29, 2022Updated 3 years ago
- TLLM_QMM strips the implementation of quantized kernels of Nvidia's TensorRT-LLM, removing NVInfer dependency and exposes ease of use Pyt…☆16Jul 5, 2024Updated 2 years ago
- Replace the MS-TCN with ASFormer in asrf☆23Oct 28, 2021Updated 4 years ago