The High Performance LLM Native Mock Server
☆44Aug 4, 2026Updated last week
Alternatives and similar repositories for VidaiMock
Users that are interested in VidaiMock are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Stateless LLM runtime that dynamically routes, loads, executes, and unloads models per request with bounded VRAM caching and intelligent …☆42Apr 12, 2026Updated 4 months ago
- Open-source cross-modal and multimodal prompt injection test suite. 250,000+ attack payloads across text, image, document, and audio moda…☆68Jul 22, 2026Updated 3 weeks ago
- ☆10Jan 23, 2025Updated last year
- MockLLM, when you want it to do what you tell it to do!☆95May 29, 2026Updated 2 months ago
- win32 native frontend for llama-cli☆14Nov 2, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- finetune method to create think/model/requires tags to allow LLMs to write programs for things they can calculate instead of hallucinatin…☆15Apr 8, 2026Updated 4 months ago
- Control your AI agents remotely. Respond from anywhere.☆16Jun 8, 2026Updated 2 months ago
- Hill Space is All You Need☆17Jul 11, 2025Updated last year
- Governed memory runtime for AI assistants: policy-before-storage, context admission, memory usage trace, deletion proof, leakage evals, a…☆18Updated this week
- AI Context Takt: A pipeline tool for structural control over LLM context. Escape the black box of chat history, maximize token efficiency…☆21Jan 12, 2026Updated 7 months ago
- Drone-first photorealistic simulation and cross-benchmark trajectory tooling for VLA research☆17Apr 15, 2026Updated 4 months ago
- Network for procedural editing of text with LLMs☆23Apr 28, 2026Updated 3 months ago
- ArgosOS is an open-source system that combines file storage, tagging, and AI-driven search. Built with FastAPI and SQLite, it’s designed …☆17May 2, 2026Updated 3 months ago
- Zero-instrumentation LLM API and MCP tracer for your agents powered by eBPF — latency, tokens, and tool use in realtime☆18Mar 16, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆24Jan 22, 2025Updated last year
- A Ray Tracing-Inspired Approach to Neural Network Optimization☆17Jun 11, 2025Updated last year
- Cleanai (https://github.com/willmil11/cleanai) except I'm making it in c now. Fast and clean from the start this time :)☆15Jul 17, 2026Updated 3 weeks ago
- ☆19Jul 4, 2025Updated last year
- Run BitNet b1.58 ternary LLMs with WebGPU — in browsers and native apps☆21Aug 6, 2026Updated last week
- Local-first agent runtime for MCP workflows with explicit trust controls, replayable runs, and built-in evals.☆32Jul 4, 2026Updated last month
- ☆14Mar 8, 2025Updated last year
- Calibrating LLM Confidence by Probing Perturbed Representation Stability☆19Jul 5, 2025Updated last year
- StatelessChatUI is a lightweight, single-file, OpenAI-compatible chat frontend. It’s built to be portable and privacy-friendly: no accoun…☆19Dec 18, 2025Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Connect Chromium to LLM agents via token-efficient DOM compression. 50-200 tokens per page.☆21Updated this week
- MiRAGE: A Multiagent Framework for Generating Multimodal Multihop Question-Answer Dataset for RAG Evaluation☆22Aug 5, 2026Updated last week
- ☆19Oct 25, 2025Updated 9 months ago
- ☆21Jul 25, 2025Updated last year
- Profiling Google Gemma 3n Model Using PyTorch Profiler☆17Jul 7, 2025Updated last year
- Bloat Free, Portable and Lightweight LLM Frontend (Single HTML file). With Lorebook, Web Search, Macro Engine etc.☆22Aug 1, 2026Updated 2 weeks ago
- ☆20Nov 26, 2025Updated 8 months ago
- A PyTorch implementation of gradient-free optimization for directly optimizing NDCG (Normalized Discounted Cumulative Gain) in neural inf…☆19Dec 21, 2025Updated 7 months ago
- Live, zero-config Redis traffic profiler built on eBPF. Reads plaintext and TLS, no app changes.☆25Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- One-Click RAG Implementation, Simple and Portable☆30Oct 5, 2025Updated 10 months ago
- A minimal, expressive, domain-specific language (DSL) designed for ultra-dense knowledge encoding in LLM prompts.☆20Feb 27, 2026Updated 5 months ago
- This repo provides a simple Gradio UI to run Qwen2 VL 72B AWQ in venv and have both image and video inferencing work.☆33Oct 3, 2024Updated last year
- The GPU-free LLM inference engine. Combines lazy expert loading + TurboQuant KV compression to run models that shouldn't fit on your hard…☆23Apr 13, 2026Updated 4 months ago
- Node.js (ESM) service that polls Gmail, sends each new email to a local OpenAI-compatible LLM, and optionally forwards summarized mobile …☆26Feb 23, 2026Updated 5 months ago
- Cursor system prompt repository☆28Nov 25, 2024Updated last year
- A Singer tap that wraps Airbyte sources allowing them to be consumed by Singer targets☆26Mar 20, 2025Updated last year