☆55Mar 22, 2025Updated last year
Alternatives and similar repositories for Bonsai
Users that are interested in Bonsai are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 33B Chinese LLM, DPO QLORA, 100K context, AirLLM 70B inference with single 4GB GPU☆14May 5, 2024Updated 2 years ago
- LLamaHTML is a simple html file to communicate with a running llamacpp llama-server☆25Aug 5, 2025Updated last year
- ☆16Feb 1, 2025Updated last year
- A sleek, customizable interface for managing LLMs with responsive design and easy agent personalization.☆19Aug 30, 2024Updated last year
- a character-ai like UI for LLM☆10Dec 3, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆87Jun 20, 2025Updated last year
- Holly - host your own AI coding agent inside docker container. Keep your system safe.☆23Jul 13, 2026Updated 3 weeks ago
- A forward proxy to turn network traffic into personal memory for AI agents☆38Mar 30, 2026Updated 4 months ago
- ☆17Mar 20, 2026Updated 4 months ago
- ☆93Sep 30, 2024Updated last year
- 🤖 AI-powered CLI for file reorganization. Runs fully locally — no data leaves your machine.☆20Jul 2, 2025Updated last year
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech G…☆26Mar 28, 2025Updated last year
- [KDD 2023] code for "Test accuracy vs. generalization gap: model selection in NLP without accessing training or testing data" https://arx…☆12Oct 17, 2022Updated 3 years ago
- A miniaturized version of the Kimi-K2 model optimized for deployment on single H100 GPUs.☆36Jul 16, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Testing LLM reasoning abilities with lineage relationship quizzes.☆44Mar 10, 2026Updated 4 months ago
- The reproduct of the paper - Aligner: Achieving Efficient Alignment through Weak-to-Strong Correction☆21May 29, 2024Updated 2 years ago
- ☆24Jan 22, 2025Updated last year
- Tcurtsni: Reverse Instruction Chat, ever wonder what your LLM wants to ask you?☆23Jun 25, 2024Updated 2 years ago
- Implementation and explorations into Blackbox Gradient Sensing (BGS), an evolutionary strategies approach proposed in a Google Deepmind p…☆20Apr 17, 2026Updated 3 months ago
- Anthropic's Contextual Retrieval implementation with visual chunk comparison. Preview context enrichment before/after embedding.☆30Sep 25, 2025Updated 10 months ago
- Inspect LLM's logprobs and perplexity over a piece of text, or compare two LLMs (like a git diff)☆19Updated this week
- Translate your large PDF files with ease using the power of Large Language Models!☆23Sep 8, 2025Updated 11 months ago
- ☆16Mar 10, 2026Updated 5 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- wagwan my slime☆12Nov 24, 2020Updated 5 years ago
- My version of an LLM Websearch Agent using a local SearXNG server because SearXNG is great.☆47Jan 27, 2026Updated 6 months ago
- Compositional Muon release☆23Jun 5, 2026Updated 2 months ago
- npm package template with typescript and tsup☆11Nov 27, 2025Updated 8 months ago
- ☆15May 2, 2026Updated 3 months ago
- EvaByte: Efficient Byte-level Language Models at Scale☆119Apr 22, 2025Updated last year
- For the better CI as well as CD using gogs and drone base on kubernetes☆10Jul 31, 2021Updated 5 years ago
- RAIL wireless applications. Go to https://github.com/SiliconLabs/application_examples☆13Jul 24, 2026Updated 2 weeks ago
- Reference implementation of "Softmax Attention with Constant Cost per Token" (Heinsen, 2024)☆25Jun 6, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- My Implementation of Q-Sparse: All Large Language Models can be Fully Sparsely-Activated☆37Aug 14, 2024Updated last year
- Context Hub Runtime Environment (CHRE)☆11Mar 13, 2026Updated 4 months ago
- Heterogeneous Scientific Foundation Model Collaboration☆23May 1, 2026Updated 3 months ago
- A bot that checks your grammar and phrasing using LLM of choice☆35Feb 6, 2025Updated last year
- Serve local ML inference engines to web apps☆32Apr 9, 2024Updated 2 years ago
- ☆11Feb 28, 2022Updated 4 years ago
- Vite utility for vue3 server side rendering☆10Jul 21, 2026Updated 2 weeks ago