Give text-only LLMs vision. A tiny OpenAI-compatible proxy that lets reasoning models (DeepSeek, Qwen, GLM…) see images by querying a separate vision model through tools: look, OCR, scan, crop, compare. No training, no weights.
☆49Jul 7, 2026Updated last month
Alternatives and similar repositories for visionbridge
Users that are interested in visionbridge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A powerful MCP testing tool with multi-provider LLM support (Ollama, OpenAI, Claude, Gemini). Test, debug, and develop MCP servers with a…☆18Apr 28, 2026Updated 3 months ago
- A zero-dependency Python module for inspecting and converting coding-agent session files (.jsonl) — Claude Code, Codex, and Pi — into the…☆47Jul 10, 2026Updated last month
- Qwen3.5-122B-A10B on a DGX Spark with DFlash speculative decode. One-shot Docker/vLLM installer. 80+ tok/s!☆60Jun 29, 2026Updated last month
- Self-hosted voice for coding agents. Talk from any browser or a Telegram call, interrupt mid-sentence, clone any voice, and hand real wor…☆36Aug 17, 2026Updated last week
- An Enhanced TOP program to monitor your Nvidia DGX SPARK's Hardware☆35Jan 6, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A pi extension that adds a /for-$each prompt loop – and hides it from your LLM! Supports children-in-directory and line-in-files iteratio…☆18Jul 20, 2026Updated last month
- Deploy DeepSeek V4 Flash (MoE reasoning model) on dual DGX Spark nodes with 1M token context, InfiniBand, and FP8 KV-cache☆103Updated this week
- Use claude code with openrouter or any other openai compatible endpoint☆23Jul 23, 2025Updated last year
- ☆52Dec 4, 2025Updated 8 months ago
- FlashQLA TileLang GDN kernels ported to NVIDIA Blackwell consumer (GB10 / DGX Spark)☆17Jun 5, 2026Updated 2 months ago
- Reactive AI - RxLM: Reactive Language Models - training and inference framework. Part of RxNN Platform Ecosystem. Licensed under custom "…☆25May 27, 2026Updated 2 months ago
- VellumForge2 is a Golang CLI for generating high-quality Direct Preference Optimization datasets via a hierarchical prompt pipeline with …☆21Feb 15, 2026Updated 6 months ago
- ☆17Aug 13, 2026Updated last week
- Auger forwards TCP traffic from localhost to a server☆15May 31, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Deep Learning technology to upscale music.☆23Jun 17, 2020Updated 6 years ago
- A local-first, privacy-preserving alternative to Google NotebookLM — Tauri 2, React 19, Rust, Ollama☆44Jul 30, 2026Updated 3 weeks ago
- TRELLIS.2 image-to-3D in C++/GGML (CUDA + Vulkan), with a resident HTTP server☆262Updated this week
- Official repository of "Efficient and Effective Query Expansion for Web Search", Short Paper @ CIKM 2018☆15Nov 17, 2019Updated 6 years ago
- Web UI for hermes, PWA installable app, extracted from hermes desktop, mobile app available.☆99Updated this week
- bf16 LoRA fine-tuning of [Qwen3.5-35B-A3B](https://huggingface.co/unsloth/Qwen3.5-35B-A3B) (a 35B-total / 3B-active Mixture-of-Experts vi…☆17Mar 12, 2026Updated 5 months ago
- Native MLX quants of Qwen3.6-27B-AEON-Ultimate-Uncensored for Apple Silicon (Metal): 8-bit, FP4, and MTP self-speculation. Preserves the …☆42Jun 28, 2026Updated last month
- MCP proxy that reduces token usage via TOON format☆19Feb 28, 2026Updated 5 months ago
- Run Pi coding agent isolated in a Docker Sandbox microVM with a local llama-server as the inference backend☆27Jun 16, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Interactive Jacobian-Lens visualizer and live steerer for GGUF models on llama.cpp☆83Jul 12, 2026Updated last month
- A Prometheus metrics exporter for NVIDIA DGX Spark clusters.☆21Feb 16, 2026Updated 6 months ago
- A full GUI experience on top of llama.cpp: all-knobs model tuning, one-click build/update from upstream, HuggingFace discovery with VRAM-…☆63Aug 9, 2026Updated 2 weeks ago
- ☆15Jul 2, 2026Updated last month
- CLI for spins up a K8s cluster locally in 10 seconds.☆16Jun 4, 2024Updated 2 years ago
- Requires Unreal Engine 4.20☆11Mar 11, 2020Updated 6 years ago
- Web UI for sparkrun — launch and monitor inference workloads on NVIDIA DGX Spark☆25Jun 16, 2026Updated 2 months ago
- Hierarchical RAG architecture scaling to 693K chunks on consumer hardware (4GB VRAM). Features 3-address routing, hybrid vector+graph fus…☆39Feb 11, 2026Updated 6 months ago
- A game for GameOff 2018☆15Dec 1, 2018Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Hermes Relay for Chrome gives Hermes Agent a direct browser surface for page context, capture, watchlists, and AI handoff.☆22May 4, 2026Updated 3 months ago
- Open-source self-hosted home AI inference platform for AMD Strix Halo — multi-backend slots, OpenAI-compatible gateway, Vue 3 + FastAPI +…☆68Updated this week
- DGX Spark inference dashboard — vLLM, SGLang, llama.cpp, WebGPU & sparkrun with agent-powered Auto-Fix and speed optimization☆57Aug 3, 2026Updated 2 weeks ago
- Working recipe to serve DeepSeek-V4-Flash across two NVIDIA DGX Spark (GB10) nodes with vLLM (TP=2, FP8 KV, MTP) over a RoCE/RDMA link — …☆71Jul 13, 2026Updated last month
- Ansible All The Things!☆14Apr 30, 2026Updated 3 months ago
- ☆11Mar 25, 2019Updated 7 years ago
- ☆10Nov 11, 2019Updated 6 years ago