Inference Model Manager for Kubernetes
☆46Apr 10, 2019Updated 7 years ago
Alternatives and similar repositories for inference-model-manager
Users that are interested in inference-model-manager are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A scalable inference server for models optimized with OpenVINO™☆921Updated this week
- A multi-user, distributed computing environment for running DL model training experiments on Intel® Xeon® Scalable processor-based system…☆390May 10, 2024Updated 2 years ago
- Bridge to connect nGraph with TensorFlow☆52Jan 3, 2023Updated 3 years ago
- Intel® AI Reference Models: contains Intel optimizations for running deep learning workloads on Intel® Xeon® Scalable processors and Inte…☆731Feb 11, 2026Updated 6 months ago
- ☆11Nov 20, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18May 29, 2026Updated 3 months ago
- IoT JavaScript API specifications and test suites☆19Aug 4, 2022Updated 4 years ago
- A demo to showcase the power of ServiceWorker and the Physical Web☆17Apr 26, 2015Updated 11 years ago
- Python bindings for OpenSHMEM☆27Updated this week
- 根据夏曹俊老师的课程,整理出来的demo☆10Aug 19, 2019Updated 7 years ago
- Xfce Desktop container designed for direct access to the GPU with EGL using VirtualGL for GPUs. Does not require /tmp/.X11-unix host sock…☆10Jul 25, 2022Updated 4 years ago
- C++ code and documentation for the MFlash PKDD'16 publication☆10Oct 25, 2016Updated 9 years ago
- Mathematical expression evaluator with just in time code generation.☆12Apr 7, 2013Updated 13 years ago
- Commandline tools for the Opus audio codec☆13Jun 24, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Distributed SDDMM Kernel☆13Jul 8, 2022Updated 4 years ago
- Provides an RPC interface to automate VSCode from other processes☆11May 29, 2021Updated 5 years ago
- ☆15Mar 7, 2018Updated 8 years ago
- Nonblocking data structures☆12Jan 25, 2015Updated 11 years ago
- Thrift SASL module that implements TSaslClientTransport☆18May 26, 2021Updated 5 years ago
- CUDA C simple application for Nvidia's GPU☆11Jun 7, 2022Updated 4 years ago
- libpypa is a Python parser implemented in pure C++☆10May 10, 2015Updated 11 years ago
- Skills for RAGFlow dataset api☆17Mar 24, 2026Updated 5 months ago
- Distributed Deep Learning Benchmark Suite☆11Oct 31, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Spring MVC Form Handling Example☆13May 7, 2019Updated 7 years ago
- TensorFlow plugin for Gen probabilistic programming system.☆10Apr 7, 2021Updated 5 years ago
- ☆12Jun 3, 2019Updated 7 years ago
- Math-aware QA system☆18May 8, 2026Updated 3 months ago
- This project provides various tools for processing content MathML with Java.☆13Feb 17, 2026Updated 6 months ago
- The socket.io layer of Overleaf for real-time editor interactions☆17Aug 6, 2021Updated 5 years ago
- Wind River Linux Setup -- Distribution Build Project Assembler☆12Oct 2, 2019Updated 6 years ago
- Example of applying CUDA graphs to LLaMA-v2☆11Aug 25, 2023Updated 3 years ago
- Catamount is a compute graph analysis tool to load, construct, and modify deep learning models and to symbolically analyze their compute …☆14May 18, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- analytics tool kit☆41Jan 23, 2017Updated 9 years ago
- A simple tool for parsing the profile.json file of mxnet☆14Aug 1, 2018Updated 8 years ago
- Simple starter CMake project that uses NVBench.☆15May 6, 2025Updated last year
- Baidu Hook☆13Jan 7, 2016Updated 10 years ago
- Web version of “Neuroevolution of Self-Interpretable Agents” (https://arxiv.org/abs/2003.08165)☆22Jan 12, 2022Updated 4 years ago
- SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, …☆2,706Updated this week
- Inline PTX Assembly in CUDA example☆15May 7, 2022Updated 4 years ago