LLM Semantic Router: Intelligent Mixture-of-Models (MoM) System with Privacy Preservation and Prompt Guard. The semantic router intelligently directs OpenAI compliant API requests to the most suitable backend models based on semantic understanding of request content.
☆22Aug 30, 2025Updated last year
Alternatives and similar repositories for semantic_router
Users that are interested in semantic_router are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Proxy-Wasm module allowing communication to Authorino and Limitador.☆15Updated this week
- Kube State Metrics `CustomResourceState` configurations for Gateway API resources☆28Aug 24, 2026Updated last month
- Use IBM Granite LLM as your Code Assistant in Visual Studio Code☆16Mar 14, 2025Updated last year
- GITHUB TEMPLATE — Click "Use this template" above or see link below for docs:☆15Aug 15, 2026Updated last month
- A demo helps you have a quick start to Tencent Cloud Mesh 🚀☆11Sep 9, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repository contains resources, documentation and artifacts describing LLM agents☆15Jan 22, 2025Updated last year
- BPF examples for Kubernetes☆14May 25, 2019Updated 7 years ago
- ☆19Jul 24, 2026Updated 2 months ago
- ☆11Jul 14, 2017Updated 9 years ago
- Model Server for Kepler☆29Mar 19, 2026Updated 6 months ago
- Platform for analyzing and recommending Python packages and Python software stacks not only for AI/ML applications☆17Jun 29, 2020Updated 6 years ago
- Image pre-caching controller service☆34Oct 17, 2015Updated 10 years ago
- Visit the MLOps Guide:☆12Sep 18, 2025Updated last year
- Enhanced MCP Server: Advanced Tool for LLM Sequential Thinking☆17May 25, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The living Trust and Safety User Guide for the AI Alliance (https://thealliance.ai)☆16Aug 11, 2026Updated last month
- Model as a Service☆45Updated this week
- Vision/Mission https://ambient-code.ai : Virtual team management and collaboration platform. User guides: https://ambient-code.github.io/…☆131Updated this week
- Examples for building and running LLM services and applications locally with Podman☆208Feb 13, 2026Updated 7 months ago
- AIOps for Distributed Environments☆17Updated this week
- ☆15May 28, 2024Updated 2 years ago
- ☆22Jul 8, 2026Updated 3 months ago
- The Z Toolkit is a VS Code extension that enables developers to quickly configure their HP Z devices for model fine-tuning and local infe…☆20Sep 4, 2026Updated last month
- 大批量稳定的获取ChatGPT登录的AccessToken,同时增加静态住宅IP代理池。保证批量启动GPT-4服务时,AccessToken增加缓存时间和过期时间校验。☆14Jul 5, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Red Hat API Integration & Management Workshop☆15Jan 28, 2022Updated 4 years ago
- This is a OpenMetadata custom connector to any spatial data format which can be read through fiona (the OGR part of the excellent GDAL li…☆22Sep 12, 2024Updated 2 years ago
- Open Enterprise Agent Governance & Orchestration☆21Sep 8, 2026Updated last month
- Automatic peer management for Consul in Kubernetes☆22Jan 21, 2021Updated 5 years ago
- ☆14Dec 8, 2021Updated 4 years ago
- Run Knative on Raspberry Pi in 5 minutes☆32Nov 14, 2021Updated 4 years ago
- Best practice website☆14Sep 10, 2025Updated last year
- ☆18Jun 4, 2020Updated 6 years ago
- ☆21Jan 29, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- An external provider for Llama Stack allowing for the use of RamaLama for inference.☆22Dec 22, 2025Updated 9 months ago
- ☆32Nov 5, 2016Updated 9 years ago
- sigstore maven plugin☆19Jul 22, 2024Updated 2 years ago
- Cython bindings for Turbo Base64☆15Aug 7, 2023Updated 3 years ago
- Experiment with cgroup-ebpf☆17Jan 13, 2019Updated 7 years ago
- ☆10Feb 18, 2024Updated 2 years ago
- Carbon Limiting Auto Tuning for Kubernetes☆38Mar 19, 2026Updated 6 months ago