π LLM Context Benchmarks - A comprehensive benchmarking tool for testing LLMs with varying context sizes using Ollama. Features dual benchmark modes (API/CLI), automatic hardware detection (optimized for Apple Silicon), visual performance charts.
β78Jul 24, 2026Updated this week
Alternatives and similar repositories for llm_context_benchmarks
Users that are interested in llm_context_benchmarks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- this repo has all official MLX-LM-LoRA example notebooks for training on Apple Siliconβ36Apr 23, 2026Updated 3 months ago
- A command-line utility to manage MLX models between your Hugging Face cache and LM Studio.β88Nov 11, 2025Updated 8 months ago
- Moshi-Finetune-MLX lets you fine-tune Moshi (Native, Real-Time, Speech-to-Speech) models all on Apple Silicon.β24Apr 21, 2026Updated 3 months ago
- Train Large Language Models on MLX.β400Updated this week
- β340May 15, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation of Visual Intelligence Using SmolVLM 2 by Hugging Faceβ39May 23, 2026Updated 2 months ago
- A proxy for minimax-m2, enabling interleaved thinking, and tool calls.β39Nov 21, 2025Updated 8 months ago
- A native Mac App for LLM fine-tuning on Apple Silicon β fully on-device, fully open source.β250Jul 16, 2026Updated last week
- A repo of useful MLX skills.β87Jan 25, 2026Updated 6 months ago
- Find the hidden meaning of LLMsβ42Nov 13, 2025Updated 8 months ago
- β34Apr 16, 2026Updated 3 months ago
- β36Mar 30, 2026Updated 3 months ago
- β30May 15, 2026Updated 2 months ago
- DeepSeek V4 Flash specific inference engine. SSD MoE expert paging (slot-bank) + disk KV cache for long agent sessions. Metal-first, narrβ¦β29Jul 18, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The best benchmark for LLMs on Apple's MLX framework knowledge and coding tasks.β36Jun 12, 2026Updated last month
- Native Mac OS GUI for Using mlx-lm-lora.β64Dec 19, 2025Updated 7 months ago
- β41Apr 9, 2026Updated 3 months ago
- Codebase exploration with AI research agentsβ20Feb 25, 2025Updated last year
- An MLX port of Meta's Coconut reasoning modelβ16Sep 2, 2025Updated 10 months ago
- REAP expert pruning for MoE LLMs on Apple Silicon via MLXβ58Mar 16, 2026Updated 4 months ago
- Pi coding agent extension that gives the agent the ability to switch models on its ownβ92Apr 14, 2026Updated 3 months ago
- The ultimate training toolkit for finetuning diffusion modelsβ34Jan 22, 2026Updated 6 months ago
- Test LLMs on real tasks. Compare models side-by-side.β382Jun 16, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- mactop - Apple Silicon Monitor Topβ1,516Jun 25, 2026Updated last month
- This repo maintains a 'cheat sheet' for LLMs that are undertrained on mlxβ33Mar 12, 2026Updated 4 months ago
- MLX Model Quantization Toolkit - Comprehensive collection of Jupyter notebooks for converting and quantizing large language models usiβ¦β16Aug 16, 2025Updated 11 months ago
- β28Mar 11, 2026Updated 4 months ago
- Automated hyperparameter search for optimal Gabliteration configurations on large language modelsβ49Mar 10, 2026Updated 4 months ago
- β18May 27, 2025Updated last year
- Multi-agent collaboration extension for pi β task decomposition, dependency management, parallel execution, TUI panelβ47Updated this week
- Harbor agent adapter for pi coding agent to run Terminal-Bench evaluationsβ31Dec 1, 2025Updated 7 months ago
- β16Feb 21, 2026Updated 5 months ago
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- QMC Heston pricing on GPUβ12May 2, 2023Updated 3 years ago
- Evaluating practical performance of local multi-turn conversational LLMs.β19Aug 1, 2025Updated 11 months ago
- β204Jun 11, 2026Updated last month
- β36Apr 25, 2026Updated 3 months ago
- Examples on how to use various LLM providers with a Wine Classification problemβ129Apr 21, 2026Updated 3 months ago
- DFlash block-diffusion speculative decoding running on Apple Silicon via MLX, with an ANE execution path that explores heterogeneous acceβ¦β62Apr 18, 2026Updated 3 months ago
- Open source, clear, transparent, real world llm benchmarksβ49Updated this week