☆42Apr 16, 2026Updated 3 months ago
Alternatives and similar repositories for kvc-3fs-operator
Users that are interested in kvc-3fs-operator are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A workload for deploying LLM inference services on Kubernetes☆262Updated this week
- Historical Kubernetes evidence for humans and agents.☆18Updated this week
- Kubernetes controllers for fast model actuation using vLLM sleep/wake and launcher-based model swapping☆16Updated this week
- k8s CSI driver for FastCFS☆13Mar 17, 2024Updated 2 years ago
- ☆10Feb 17, 2026Updated 5 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- m3fs(Make 3FS) is the toolset designed to deploy 3FS cluster.☆62Jan 16, 2026Updated 6 months ago
- ☆25Jul 13, 2021Updated 5 years ago
- FalconFS is a high-performance distributed file system (DFS) designed for AI workloads.☆65May 19, 2026Updated 2 months ago
- ☆18Mar 1, 2025Updated last year
- A deploy bot for personal use.☆13Mar 31, 2019Updated 7 years ago
- A benchmark suite for evaluating FaaS scheduler.☆23Nov 5, 2022Updated 3 years ago
- Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, T…☆481Updated this week
- A high-performance DNS Server based on DPDK☆23Jun 12, 2020Updated 6 years ago
- This is the final project of 2020 DBMS course in SYSU☆10Jun 23, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Alibaba Cloud's high-performance KVCache system for LLM inference, with components for global cache management, inference simulation(HiSi…☆215Updated this week
- Source code for Jellyfish, a soft real-time inference serving system☆15Dec 20, 2022Updated 3 years ago
- We present a set of all-reduce compatible gradient compression algorithms which significantly reduce the communication overhead while mai…☆10Nov 14, 2021Updated 4 years ago
- Medusa: Accelerating Serverless LLM Inference with Materialization [ASPLOS'25]☆12Nov 8, 2024Updated last year
- Ruta demo☆37Feb 18, 2021Updated 5 years ago
- [NeurIPS 2024] Implementation of paper - D-LLM: A Token Adaptive Computing Resource Allocation Strategy for Large Language Models☆24Apr 9, 2025Updated last year
- SocksDirect code repository☆20May 6, 2026Updated 2 months ago
- Operator for web-ui☆12Nov 25, 2024Updated last year
- ☆13Jun 29, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- DRAM/SSD hybrid caching system☆15Mar 13, 2025Updated last year
- ☆13Dec 18, 2020Updated 5 years ago
- etcd 注解版☆17Dec 7, 2022Updated 3 years ago
- ☆17May 10, 2024Updated 2 years ago
- Bypassd is a novel I/O architecture that provides low latency access to shared SSDs.☆23May 14, 2025Updated last year
- ☸️ Easy, advanced inference platform for large language models on Kubernetes. 🌟 Star to support our work!☆309Jan 26, 2026Updated 5 months ago
- Simple server that scrapes Predixy stats and exports them via HTTP for Prometheus consumption☆13Dec 11, 2019Updated 6 years ago
- ☆21May 27, 2025Updated last year
- d.run website☆17Jul 3, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The main purpose of runtime copilot is to assist with node runtime management tasks such as configuring registries, upgrading versions, i…☆13May 16, 2023Updated 3 years ago
- Furnace is a high-performance quantitative trading library that provides features similar to CCXT, allowing developers to connect and int…☆16Jan 16, 2025Updated last year
- ☆15Nov 17, 2015Updated 10 years ago
- LoRAFusion: Efficient LoRA Fine-Tuning for LLMs☆28Jul 2, 2026Updated 2 weeks ago
- A QA system based on k8s-specific knowledge build on ChatGLM2-6B, serving by Ray.☆10Sep 14, 2023Updated 2 years ago
- kstone-etcd-operator is a subproject of etcd cluster management platform kstone. It's inspired by etcd-operator. And has more complete su…☆20Feb 20, 2022Updated 4 years ago
- ☆13Jun 17, 2019Updated 7 years ago