A char level language model ,this repo is just for learning .
☆18Jun 14, 2026Updated last month
Alternatives and similar repositories for minilm
Users that are interested in minilm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Simple agent framework using Ollama tool calling☆10Aug 27, 2024Updated last year
- Load and run Llama from safetensors files in C☆16Oct 24, 2024Updated last year
- Implementation of build an LLM from scratch by Sebastian Raschka.☆19Jan 22, 2026Updated 6 months ago
- world's stupidest moe llm in 103M parameters☆20Jul 18, 2025Updated last year
- Ultra-lightweight C++ inference engine for BitMamba-2 (1.58-bit SSM). Runs 1B models on consumer CPUs at 50+ tok/s using <700MB RAM. No h…☆21Jun 2, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An analysis of the released data on FinCrime Files transactions as depicted on SARs.☆10Nov 26, 2020Updated 5 years ago
- ☆22Mar 18, 2026Updated 4 months ago
- An experimental desktop client for using Claude Desktop's MCP with Novelcrafter codices.☆11Dec 3, 2024Updated last year
- ☆12Jan 19, 2024Updated 2 years ago
- Red by Example - an accessible reference by example☆14Nov 29, 2022Updated 3 years ago
- Prefixed unique ID field for Django models☆15May 9, 2026Updated 2 months ago
- SwiftLet is a lightweight Python framework for running open-source Large Language Models (LLMs) locally using safetensors☆29Aug 6, 2025Updated 11 months ago
- Headless front-end for ecommerce B2B supporting different platforms and open source components☆13Jan 22, 2024Updated 2 years ago
- ☆16Mar 2, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Lightweight C inference for Qwen3 GGUF. Multiturn prefix caching & batch processing.☆25Sep 1, 2025Updated 11 months ago
- Summaries of AI Research Papers☆18Jun 29, 2024Updated 2 years ago
- Governed memory runtime for AI assistants: policy-before-storage, context admission, memory usage trace, deletion proof, leakage evals, a…☆15Updated this week
- Experience the power of AI with this free AI voice generator demo. Utilizing Deepgram and Groq, we transform text into voice seamlessly. …☆38Jun 12, 2024Updated 2 years ago
- This repository contains the code for implementation of RAG approach with company policies data, evaluation of RAG solution and smart chu…☆16Sep 18, 2025Updated 10 months ago
- AI agent memory you can trust☆15Jul 14, 2026Updated 2 weeks ago
- Evolutionary Model trained for MNIST☆16Feb 8, 2026Updated 5 months ago
- ☆17Mar 27, 2026Updated 4 months ago
- ☆18Jun 15, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- jQuery, React and Streamlit applications written by LLMs☆16Dec 24, 2023Updated 2 years ago
- ☆48Mar 9, 2025Updated last year
- Group-relative Trajectory-based Policy Optimization: Increasing Quality and Training Stability☆42Feb 23, 2026Updated 5 months ago
- An educational Rust project for exporting and running inference on Qwen3 LLM family☆43Aug 3, 2025Updated last year
- Pure Markdown rendered as HTML in the browser☆13Mar 10, 2025Updated last year
- llama.cpp fork with AMD XDNA2 NPU backend for Ryzen AI MAX (npu5/XDNA2) — matrix multiply offload via XRT☆19Apr 1, 2026Updated 4 months ago
- Writing Extension for Text Generation WebUI☆67Aug 7, 2025Updated 11 months ago
- Python AI development environment☆20Jul 17, 2025Updated last year
- replacement of AdamW and Lion optimizer for LLMs☆13May 28, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆46Jan 8, 2025Updated last year
- LaSearch releases☆19Jul 8, 2026Updated 3 weeks ago
- JavaScript bindings for the ggml-js library☆44Nov 10, 2025Updated 8 months ago
- Golang web client for Ollama, fast and easy to use.☆33Jul 18, 2025Updated last year
- 🛰 Uplink cluster management for Instellar☆13Jun 12, 2026Updated last month
- A Mix formatter for your javascript and css files using Prettier☆10Dec 11, 2025Updated 7 months ago
- Call your local Claude Code sessions. Literally, from a phone call.☆19Apr 7, 2026Updated 3 months ago