A vLLM plugin built on the FlagOS unified multi-chip backend.
☆88Sep 2, 2026Updated this week
Alternatives and similar repositories for vllm-plugin-FL
Users that are interested in vllm-plugin-FL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FlagCX is a scalable and adaptive cross-chip communication library.☆225Updated this week
- FlagScale is a large model toolkit based on open-sourced projects.☆533Updated this week
- FlagGems is an operator library for large language models implemented in the Triton Language.☆1,089Updated this week
- Next-Generation AI-Assisted Kernel Engineering for Multi-Chip Systems☆75Updated this week
- FlagTree is a unified compiler supporting multiple AI chip backends for custom Deep Learning operations, which is forked from triton-lang…☆324Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- FlagOS skills for model deployment, HW adaptation, train&infer, eval, kernel dev and perf tuning☆19Jul 18, 2026Updated last month
- A Flexible Framework for Comprehensive Multimodal Model Evaluation☆107Apr 21, 2026Updated 4 months ago
- ☆19Mar 4, 2025Updated last year
- 算法可视化模型,基于已有的模型做可视化推理,帮助更好的理解☆25Updated this week
- A Visual Studio project demonstrating how to perform object tracking across video frames with YOLOX, ONNX Runtime, and the ByteTrack-Eige…☆12Nov 19, 2023Updated 2 years ago
- "Personal site documenting my journey and idae in CV and SLAM.☆16May 30, 2026Updated 3 months ago
- ☆13Sep 8, 2024Updated last year
- 先进编译实验室的个人主页☆254Oct 15, 2025Updated 10 months ago
- Equal Loudness Filter☆11Mar 4, 2019Updated 7 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Audio Loudness Normalization Filter Port From FFmpeg☆12Mar 4, 2019Updated 7 years ago
- 天池Better Synth多模态大模型数据合成挑战赛-打赢baseline就算成功方案☆30Oct 30, 2025Updated 10 months ago
- ☆17May 1, 2024Updated 2 years ago
- Official implementation of Acc-SpMM: Accelerating General-purpose Sparse Matrix-Matrix Multiplication with GPU Tensor Cores.☆38Nov 13, 2025Updated 9 months ago
- 国科大研究生课程 操作系统高级教程2023年思考题☆12Dec 24, 2023Updated 2 years ago
- demos using speex☆12Apr 20, 2018Updated 8 years ago
- ☆16Oct 18, 2021Updated 4 years ago
- Huawei Ascend Mate 7 kernel tree☆12Sep 20, 2016Updated 9 years ago
- A skill for automatically optimizing CUDA code.☆43Mar 26, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆15Sep 19, 2018Updated 7 years ago
- 适用于sophon bm1684x,基于 Langchain 与 ChatGLM 等语言模型的本地知识库问答☆13Jun 5, 2024Updated 2 years ago
- This repository is webrtc agc module demo.☆12Jan 23, 2019Updated 7 years ago
- [QReward] RewardService Python Client, make RL Training reward function more faster☆16Apr 23, 2026Updated 4 months ago
- Community maintained hardware plugin for vLLM on Ascend☆2,748Updated this week
- ☆14Nov 5, 2025Updated 9 months ago
- ☆68Jul 14, 2025Updated last year
- Venus Collective Communication Library, supported by SII and Infrawaves.☆151Jun 24, 2026Updated 2 months ago
- ☆27Mar 31, 2026Updated 5 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆16Apr 3, 2025Updated last year
- Some DSP algorithm implementation☆18Sep 26, 2018Updated 7 years ago
- SGLang is a fast serving framework for large language models and vision language models.☆32Updated this week
- An interface to program any congestion control protocol for an unreliable connection based protocol sent over UDP. It comes with a clean …☆12Apr 8, 2022Updated 4 years ago
- webRTCtest-linux☆12Jun 21, 2018Updated 8 years ago
- 多级缓存架构项目,kafka读取 缓存更新请求,从数据库获取缓存,写入ehcache和redis☆11Jul 6, 2018Updated 8 years ago
- AI语音输入法。高性能、低延迟、易配置,你的数据永远在本地。☆16Feb 22, 2026Updated 6 months ago