π LLM Context Benchmarks - A comprehensive benchmarking tool for testing LLMs with varying context sizes using Ollama. Features dual benchmark modes (API/CLI), automatic hardware detection (optimized for Apple Silicon), visual performance charts.
β97Sep 19, 2026Updated this week
Alternatives and similar repositories for llm_context_benchmarks
Users that are interested in llm_context_benchmarks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- this repo has all official MLX-LM-LoRA example notebooks for training on Apple Siliconβ42Sep 13, 2026Updated last week
- Moshi-Finetune-MLX lets you fine-tune Moshi (Native, Real-Time, Speech-to-Speech) models all on Apple Silicon.β26Apr 21, 2026Updated 5 months ago
- Train Large Language Models on MLX.β418Updated this week
- β341May 15, 2026Updated 4 months ago
- Implementation of Visual Intelligence Using SmolVLM 2 by Hugging Faceβ39Aug 31, 2026Updated 3 weeks ago
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A proxy for minimax-m2, enabling interleaved thinking, and tool calls.β38Nov 21, 2025Updated 10 months ago
- A native Mac App for LLM fine-tuning on Apple Silicon β fully on-device, fully open source.β269Aug 26, 2026Updated 3 weeks ago
- Pi extension that tracks bash tool token usage with live stats, grouping, and exportβ23Feb 10, 2026Updated 7 months ago
- A repo of useful MLX skills.β89Jan 25, 2026Updated 7 months ago
- Find the hidden meaning of LLMsβ42Nov 13, 2025Updated 10 months ago
- β34Apr 16, 2026Updated 5 months ago
- β38Mar 30, 2026Updated 5 months ago
- β32May 15, 2026Updated 4 months ago
- DeepSeek V4 Flash specific inference engine. SSD MoE expert paging (slot-bank) + disk KV cache for long agent sessions. Metal-first, narrβ¦β31Sep 6, 2026Updated 2 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The best benchmark for LLMs on Apple's MLX framework knowledge and coding tasks.β38Jun 12, 2026Updated 3 months ago
- Native Mac OS GUI for Using mlx-lm-lora.β66Dec 19, 2025Updated 9 months ago
- An Enhanced TOP program to monitor your Nvidia DGX SPARK's Hardwareβ37Jan 6, 2026Updated 8 months ago
- β41Apr 9, 2026Updated 5 months ago
- Tritonβstyle kernel toolkit for MLX plus a small upstream incubator: prototype, benchmark, and upstream fusions for Apple Siliconβ53Mar 31, 2026Updated 5 months ago
- Codebase exploration with AI research agentsβ21Feb 25, 2025Updated last year
- An MLX port of Meta's Coconut reasoning modelβ16Sep 2, 2025Updated last year
- A tutorial on Apple MLX and LLMβ61May 16, 2025Updated last year
- Universal Ethernet Direct Connect (UniEDC)β32Apr 1, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Run ops on Apple ANE in NPU register with pure python on M1 Asahi Linux. No Espresso, No CoreML, no metal, no .mlmodels file, no .hwx filβ¦β19Jun 28, 2026Updated 2 months ago
- REAP expert pruning for MoE LLMs on Apple Silicon via MLXβ58Mar 16, 2026Updated 6 months ago
- Pi coding agent extension that gives the agent the ability to switch models on its ownβ101Aug 23, 2026Updated last month
- Test LLMs on real tasks. Compare models side-by-side.β423Aug 10, 2026Updated last month
- mactop - Apple Silicon Monitor Topβ1,639Sep 10, 2026Updated 2 weeks ago
- This repo maintains a 'cheat sheet' for LLMs that are undertrained on mlxβ33Mar 12, 2026Updated 6 months ago
- MLX Model Quantization Toolkit - Comprehensive collection of Jupyter notebooks for converting and quantizing large language models usiβ¦β15Aug 16, 2025Updated last year
- β29Aug 8, 2026Updated last month
- Automated hyperparameter search for optimal Gabliteration configurations on large language modelsβ49Mar 10, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- RLM (Recursive Language Model) extension for pi - process large context files that exceed LLM context windowsβ19Feb 8, 2026Updated 7 months ago
- β18May 27, 2025Updated last year
- Multi-agent collaboration extension for pi β task decomposition, dependency management, parallel execution, TUI panelβ50Aug 20, 2026Updated last month
- β16Feb 21, 2026Updated 7 months ago
- Evaluating practical performance of local multi-turn conversational LLMs.β19Aug 1, 2025Updated last year
- β12Jul 24, 2019Updated 7 years ago
- β37Apr 25, 2026Updated 4 months ago