π LLM Context Benchmarks - A comprehensive benchmarking tool for testing LLMs with varying context sizes using Ollama. Features dual benchmark modes (API/CLI), automatic hardware detection (optimized for Apple Silicon), visual performance charts.
β90Aug 30, 2026Updated this week
Alternatives and similar repositories for llm_context_benchmarks
Users that are interested in llm_context_benchmarks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A command-line utility to manage MLX models between your Hugging Face cache and LM Studio.β89Nov 11, 2025Updated 9 months ago
- Minimal Claude Code alternative powered by MLXβ47Jan 11, 2026Updated 7 months ago
- Moshi-Finetune-MLX lets you fine-tune Moshi (Native, Real-Time, Speech-to-Speech) models all on Apple Silicon.β26Apr 21, 2026Updated 4 months ago
- Train Large Language Models on MLX.β410Updated this week
- β342May 15, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of Visual Intelligence Using SmolVLM 2 by Hugging Faceβ39Updated this week
- A proxy for minimax-m2, enabling interleaved thinking, and tool calls.β39Nov 21, 2025Updated 9 months ago
- A native Mac App for LLM fine-tuning on Apple Silicon β fully on-device, fully open source.β260Aug 26, 2026Updated last week
- Repository for running LLMs efficiently on Mac silicon (M1, M2, M3). Features Jupyter notebook for Meta-Llama-3 setup using MLX frameworkβ¦β11May 4, 2024Updated 2 years ago
- Pi extension that tracks bash tool token usage with live stats, grouping, and exportβ23Feb 10, 2026Updated 6 months ago
- A repo of useful MLX skills.β89Jan 25, 2026Updated 7 months ago
- Find the hidden meaning of LLMsβ42Nov 13, 2025Updated 9 months ago
- β34Apr 16, 2026Updated 4 months ago
- β36Mar 30, 2026Updated 5 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- β32May 15, 2026Updated 3 months ago
- DeepSeek V4 Flash specific inference engine. SSD MoE expert paging (slot-bank) + disk KV cache for long agent sessions. Metal-first, narrβ¦β32Jul 18, 2026Updated last month
- The best benchmark for LLMs on Apple's MLX framework knowledge and coding tasks.β38Jun 12, 2026Updated 2 months ago
- Native Mac OS GUI for Using mlx-lm-lora.β65Dec 19, 2025Updated 8 months ago
- Chain-of-thought λ°©μμ νμ©νμ¬ llama2λ₯Ό fine-tuningβ10Nov 18, 2023Updated 2 years ago
- β41Apr 9, 2026Updated 4 months ago
- Codebase exploration with AI research agentsβ21Feb 25, 2025Updated last year
- An MLX port of Meta's Coconut reasoning modelβ16Sep 2, 2025Updated last year
- Run ops on Apple ANE in NPU register with pure python on M1 Asahi Linux. No Espresso, No CoreML, no metal, no .mlmodels file, no .hwx filβ¦β19Jun 28, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- REAP expert pruning for MoE LLMs on Apple Silicon via MLXβ58Mar 16, 2026Updated 5 months ago
- Pi coding agent extension that gives the agent the ability to switch models on its ownβ97Aug 23, 2026Updated last week
- The ultimate training toolkit for finetuning diffusion modelsβ34Jan 22, 2026Updated 7 months ago
- Test LLMs on real tasks. Compare models side-by-side.β416Aug 10, 2026Updated 3 weeks ago
- This repo maintains a 'cheat sheet' for LLMs that are undertrained on mlxβ33Mar 12, 2026Updated 5 months ago
- MLX Model Quantization Toolkit - Comprehensive collection of Jupyter notebooks for converting and quantizing large language models usiβ¦β16Aug 16, 2025Updated last year
- β29Aug 8, 2026Updated 3 weeks ago
- The SEAL-CPU backend is a Reference backend engine for HEBench which is a shared library that implements the required functions specifiedβ¦β11Mar 3, 2023Updated 3 years ago
- Incremental optimizations to the N-Body problem in order to evaluate and compare the performance of Python translators in the HPC environβ¦β13Apr 2, 2023Updated 3 years ago
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Automated hyperparameter search for optimal Gabliteration configurations on large language modelsβ49Mar 10, 2026Updated 5 months ago
- RLM (Recursive Language Model) extension for pi - process large context files that exceed LLM context windowsβ18Feb 8, 2026Updated 6 months ago
- πΎ Optimize Laravel caching with Cachetastic! Cache method results, force refresh, handle errors, and boost app performance effortlessly.β13Jan 26, 2026Updated 7 months ago
- Multi-agent collaboration extension for pi β task decomposition, dependency management, parallel execution, TUI panelβ49Aug 20, 2026Updated 2 weeks ago
- Harbor agent adapter for pi coding agent to run Terminal-Bench evaluationsβ32Dec 1, 2025Updated 9 months ago
- β16Feb 21, 2026Updated 6 months ago
- Evaluating practical performance of local multi-turn conversational LLMs.β19Aug 1, 2025Updated last year