GenAI components at micro-service level; GenAI service composer to create mega-service
☆200Sep 3, 2026Updated 2 weeks ago
Alternatives and similar repositories for GenAIComps
Users that are interested in GenAIComps are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Evaluation, benchmark, and scorecard, targeting for performance on throughput and latency, accuracy on popular evaluation harness, safety…☆41Jul 6, 2026Updated 2 months ago
- Generative AI Examples is a collection of GenAI examples such as ChatQnA, Copilot, which illustrate the pipeline capabilities of the Open…☆742Sep 13, 2026Updated last week
- Containerization and cloud native suite for OPEA☆75Jul 6, 2026Updated 2 months ago
- Easy and lightning fast training of 🤗 Transformers on Habana Gaudi processor (HPU)☆213Sep 7, 2026Updated 2 weeks ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆91Sep 2, 2026Updated 3 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- An innovative library for efficient LLM inference via low-bit quantization☆352Aug 30, 2024Updated 2 years ago
- ⚡ Build your chatbot within minutes on your favorite device; offer SOTA compression techniques for LLMs; run LLMs efficiently on Intel Pl…☆2,172Oct 8, 2024Updated last year
- Intel® AI for Enterprise Inference optimizes AI inference services on Intel hardware using Kubernetes Orchestration. It automates LLM mod…☆47Sep 8, 2026Updated 2 weeks ago
- Multi-stage LLM agent pipeline for optimizing Triton kernels on Intel XPU — from analysis to autotuning.☆21Sep 11, 2026Updated last week
- Intel® Extension for DeepSpeed* is an extension to DeepSpeed that brings feature support with SYCL kernels on Intel GPU(XPU) device. Note…☆65May 27, 2026Updated 3 months ago
- A Python package for extending the official PyTorch that can easily obtain performance on Intel platform☆2,009Mar 30, 2026Updated 5 months ago
- A repository of Dockerfiles, scripts, yaml files, Helm Charts, etc. used to build and scale the sample AI workflows with python, kubernet…☆12Feb 22, 2024Updated 2 years ago
- This repository contains Dockerfiles, scripts, yaml files, Helm charts, etc. used to scale out AI containers with versions of TensorFlow …☆79Sep 8, 2026Updated 2 weeks ago
- A simple and effective quantization toolkit for high-accuracy low-bit LLM inference|简洁且高效的量化工具包☆1,622Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Framework for enhancing LLMs for RAG tasks using fine-tuning.☆768Jun 8, 2026Updated 3 months ago
- Run Generative AI models with simple C++/Python API and using OpenVINO Runtime☆589Updated this week
- this is a repository that gives the power of mixture of workflows a concept inspired by the mixture of agents.☆13Aug 19, 2024Updated 2 years ago
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.☆14Jan 8, 2026Updated 8 months ago
- OpenAI Triton backend for Intel® GPUs☆271Updated this week
- Notes from our paper reading sessions☆16Sep 24, 2020Updated 5 years ago
- Intel® AI Builder (SuperClaw, SuperBuilder)☆249Updated this week
- Cursor for Finance☆16Aug 29, 2025Updated last year
- A list of questions that can be asked during an interview for a cloud architect position.☆11Nov 27, 2021Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆15Dec 12, 2024Updated last year
- OpenVINO LLM Benchmark☆11Dec 7, 2023Updated 2 years ago
- Multi-Modal Disease Prediction☆16Jun 27, 2024Updated 2 years ago
- Explainable AI Tooling (XAI). XAI is used to discover and explain a model's prediction in a way that is interpretable to the user. Releva…☆39Sep 22, 2025Updated last year
- Explore our open source AI portfolio! Develop, train, and deploy your AI solutions with performance- and productivity-optimized tools fro…☆78Mar 27, 2026Updated 5 months ago
- ☆61Dec 18, 2024Updated last year
- Pre-built components and code samples to help you build and deploy production-grade AI applications with the OpenVINO™ Toolkit from Intel☆217Jul 30, 2026Updated last month
- Edge Insights for Vision (eiv) is a package that helps to auto install Intel® GPU drivers and setup environment for Inference application…☆22Sep 29, 2025Updated 11 months ago
- Intel® Extension for TensorFlow*☆353Oct 29, 2025Updated 10 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Intel® System Health Inspector (aka svr-info) is a Linux command line tool used to assess the health of Intel® Xeon® processor-based serv…☆60Dec 17, 2024Updated last year
- The vLLM XPU kernels for Intel GPU☆71Updated this week
- RAGme-io is a personalized RAG agent for the web sites you visit and documents you care about☆16Sep 25, 2025Updated 11 months ago
- A dashboard for exploring timm learning rate schedulers☆20Nov 22, 2024Updated last year
- A Forth interpreter/compiler and IDE for Atari ST☆13Sep 15, 2012Updated 14 years ago
- Jan is an open source alternative to ChatGPT that runs 100% offline on your computer☆12Mar 13, 2026Updated 6 months ago
- Simplifying RAG with PostgreSQL and PGVector☆16Jul 31, 2024Updated 2 years ago