☆60Mar 22, 2025Updated last year
Alternatives and similar repositories for Bonsai
Users that are interested in Bonsai are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LLamaHTML is a simple html file to communicate with a running llamacpp llama-server☆25Aug 5, 2025Updated last year
- Yet another frontend for LLM, written using .NET and WinUI 3☆11Sep 14, 2025Updated 11 months ago
- Open Source Auth Built on Freestyle: own your auth + data https://docs.freestyle.dev/guides/authentication/☆23Jun 12, 2024Updated 2 years ago
- ☆16Feb 1, 2025Updated last year
- A sleek, customizable interface for managing LLMs with responsive design and easy agent personalization.☆19Aug 30, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Inference Llama 2 in one file of pure Haskell (A port of llama2.c from Andrej Karpathy)☆14Oct 17, 2025Updated 10 months ago
- a character-ai like UI for LLM☆10Dec 3, 2024Updated last year
- ☆92Jun 20, 2025Updated last year
- Holly - host your own AI coding agent inside docker container. Keep your system safe.☆23Jul 13, 2026Updated last month
- opensource LLM☆35Sep 20, 2025Updated 11 months ago
- A forward proxy to turn network traffic into personal memory for AI agents☆38Mar 30, 2026Updated 4 months ago
- ☆17Mar 20, 2026Updated 5 months ago
- my nix packages☆11Updated this week
- ☆14Mar 4, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Rate limiter for LLM clients☆22Jan 30, 2025Updated last year
- A miniaturized version of the Kimi-K2 model optimized for deployment on single H100 GPUs.☆36Jul 16, 2025Updated last year
- The reproduct of the paper - Aligner: Achieving Efficient Alignment through Weak-to-Strong Correction☆21May 29, 2024Updated 2 years ago
- ☆24Jan 22, 2025Updated last year
- Tcurtsni: Reverse Instruction Chat, ever wonder what your LLM wants to ask you?☆23Jun 25, 2024Updated 2 years ago
- A zero-allocation, header-only C++ BPE tokenizer for Qwen, built for maximum inference throughput.☆23Apr 3, 2026Updated 4 months ago
- Anthropic's Contextual Retrieval implementation with visual chunk comparison. Preview context enrichment before/after embedding.☆30Sep 25, 2025Updated 11 months ago
- Inspect LLM's logprobs and perplexity over a piece of text, or compare two LLMs (like a git diff)☆18Aug 3, 2026Updated 3 weeks ago
- Implementation of the fast weight product key memory from Sakana AI☆20Aug 19, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Translate your large PDF files with ease using the power of Large Language Models!☆23Sep 8, 2025Updated 11 months ago
- ☆17Mar 10, 2026Updated 5 months ago
- ☆12Dec 14, 2024Updated last year
- ☆11Sep 9, 2024Updated last year
- My version of an LLM Websearch Agent using a local SearXNG server because SearXNG is great.☆47Jan 27, 2026Updated 7 months ago
- Compositional Muon release☆25Jun 5, 2026Updated 2 months ago
- TernGEMM: General Matrix Multiply Library with Ternary Weights for Fast DNN Inference☆14Feb 22, 2022Updated 4 years ago
- ☆10Mar 13, 2017Updated 9 years ago
- EvaByte: Efficient Byte-level Language Models at Scale☆119Apr 22, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- For the better CI as well as CD using gogs and drone base on kubernetes☆10Jul 31, 2021Updated 5 years ago
- RAIL wireless applications. Go to https://github.com/SiliconLabs/application_examples☆13Jul 24, 2026Updated last month
- Multi-agent orchestration framework for AI applications - build, deploy, and manage AI agents across the full lifecycle with Forge, Conve…☆33Mar 28, 2026Updated 5 months ago
- Reference implementation of "Softmax Attention with Constant Cost per Token" (Heinsen, 2024)☆25Jun 6, 2024Updated 2 years ago
- My Implementation of Q-Sparse: All Large Language Models can be Fully Sparsely-Activated☆37Aug 14, 2024Updated 2 years ago
- Context Hub Runtime Environment (CHRE)☆11Mar 13, 2026Updated 5 months ago
- Agentic BYOK Browser-Based Website Builder☆55Updated this week