Simple model memory requirements calculator for GGUF
☆84Jan 20, 2026Updated 8 months ago
Alternatives and similar repositories for model-memory-calculator
Users that are interested in model-memory-calculator are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Kolosal AI is an OpenSource and Lightweight alternative to Ollama to run LLMs 100% offline on your device.☆15Jan 2, 2026Updated 9 months ago
- Kolosal AI is an OpenSource and Lightweight alternative to LM Studio to run LLMs 100% offline on your device.☆458May 22, 2025Updated last year
- llama.cpp's official website☆19Updated this week
- indoBERT Base-Uncased fine-tuned on Translated Squad v2.0☆19Dec 24, 2024Updated last year
- 33B Chinese LLM, DPO QLORA, 100K context, AirLLM 70B inference with single 4GB GPU☆14May 5, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Offline LLM chatbot with personalized memory — works on CPU with multi-session memory support.☆22Jan 10, 2026Updated 8 months ago
- Talk to Ava, an AI assistant running entirely in your browser. Private and no server required.☆29Jul 22, 2026Updated 2 months ago
- Serving LLMs in the HF-Transformers format via a PyFlask API☆70Sep 10, 2024Updated 2 years ago
- A user-friendly GUI for llama.cpp — convert, quantize, and run GGUF models without touching the terminal.☆25Jun 9, 2026Updated 4 months ago
- Example for a lightweight React JSON Form Builder☆14May 15, 2023Updated 3 years ago
- Chat WebUI is an easy-to-use user interface for interacting with AI, and it comes with multiple useful built-in tools such as web search …☆53Feb 10, 2026Updated 7 months ago
- ☆16Feb 1, 2025Updated last year
- A Windows tool to query various LLM AIs. Supports branched conversations, history and summaries among others.☆36May 11, 2026Updated 4 months ago
- My version of an LLM Websearch Agent using a local SearXNG server because SearXNG is great.☆48Jan 27, 2026Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- E-Wallet system uses Golang + GoFiber + MidTrans + Redis + GoCache Accompanied by the SSE server sent event system for sending notificat…☆10Feb 21, 2024Updated 2 years ago
- Quick access to any large language model from your browser.☆10Feb 16, 2026Updated 7 months ago
- A novel hybrid AI architecture leveraging Titan's-like memory and HRM-like reasoning☆27Aug 28, 2026Updated last month
- An alternate reality web browser, powered by an LLM☆19Apr 29, 2024Updated 2 years ago
- Your Python AI Coder!☆36May 21, 2025Updated last year
- ☆19Feb 23, 2026Updated 7 months ago
- ✨ Instagram using tRPC (with a NextJS backend) & React Server Components: Next Auth, Prisma & Shadcn UI.☆18Apr 8, 2025Updated last year
- Minimalistic batching application for LLMs using ASP.NET Core and LLamaSharp☆12Oct 23, 2024Updated last year
- Redact PDF/image-based documents, Word, or CSV/XLSX files using a graphical user interface. Demo: https://huggingface.co/spaces/seanpedri…☆63Sep 28, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Sentiment Analysis using Recurrent Neural Networks (RNN-LSTM) and Google News Word2Vec☆10Sep 20, 2019Updated 7 years ago
- ☆15Mar 18, 2026Updated 6 months ago
- Convert downloaded Ollama models back into their GGUF equivalent format☆99Dec 21, 2024Updated last year
- An extendable powerline plugin for clink☆10Apr 10, 2019Updated 7 years ago
- ISPConfig is a web hosting control panel for Linux servers. A shell script can be used to automate common tasks like creating email accou…☆10Aug 3, 2026Updated 2 months ago
- Backend not ready? No problem. Helix is a zero-config AI mock server that generates realistic, dynamic data for any endpoint you hit.☆25Dec 19, 2025Updated 9 months ago
- Zero-overhead, terminal-native local-LLM runtime manager. Launches, supervises, and routes local models behind one OpenAI-compatible endp…☆209Updated this week
- LlamaCards is a web application that provides a dynamic interface for interacting with LLM models in real-time. This app allows users to …☆35Aug 28, 2024Updated 2 years ago
- ☆31Sep 9, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 𝗬𝗢𝗨𝗧𝗨𝗕𝗘 𝗜𝗣 𝗕𝗔𝗡 𝗜𝗦𝗦𝗨𝗘 𝗦𝗢𝗟𝗩𝗘𝗗. 𝗠𝗨𝗦𝗜𝗖 𝗕𝗢𝗧 𝗡𝗢 🌱𝗟𝗔𝗚 𝗙𝗔𝗦𝗧 𝗦𝗣𝗘𝗘𝗗 (V2)🏵️𝗕𝗢𝗧 ʏᴛ-ᴅʟᴘ ᴇʀʀᴏʀ …☆10Jul 1, 2026Updated 3 months ago
- Serverless single HTML page access to an OpenAI API compatible Local LLM☆51Sep 9, 2025Updated last year
- Vulkan & GLSL implementation of FlashAttention-2☆15Jan 19, 2025Updated last year
- A simple Electron app to run the Ionic Creator☆15Aug 13, 2016Updated 10 years ago
- Bookmarklet to pull and run hugging face GGUF models in Ollama☆18Oct 17, 2024Updated last year
- This project implements a DNS-based service discovery mechanism using Go for Proxmox Cluster☆19Oct 17, 2024Updated last year
- MACKO: Sparse matrix vector multiplication for low sparsity☆41Apr 6, 2026Updated 6 months ago