☆86Mar 3, 2026Updated 4 months ago
Alternatives and similar repositories for ANEgpt
Users that are interested in ANEgpt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Local AI runtime for training & running small LLMs directly on Apple Neural Engine (ANE). No CoreML. No Metal. Offline, on-device fine-tu…☆110Mar 6, 2026Updated 4 months ago
- Train and run transformers directly on Apple's Neural Engine in Swift bypass coreml entirely☆157Jul 16, 2026Updated last week
- Training neural networks on Apple Neural Engine via reverse-engineered private APIs☆7,056Mar 10, 2026Updated 4 months ago
- Train Embedding Models on MLX.☆17Jun 2, 2026Updated last month
- Apple Neural Engine (ANE) LLM inference engine — reverse-engineered private APIs, Metal GPU shaders, hybrid ANE+GPU+CPU on Apple Silicon.…☆23Mar 5, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Multilingual Knowledge Graph Enhancement (EMNLP 2023)☆24Nov 28, 2023Updated 2 years ago
- An MLX port of Meta's Coconut reasoning model☆16Sep 2, 2025Updated 10 months ago
- A tool which checks compatibility of CoreML model with Apple Neural Engine☆15May 30, 2022Updated 4 years ago
- Run transformers (incl. LLMs) on the Apple Neural Engine.☆63Nov 22, 2023Updated 2 years ago
- ☆52Mar 17, 2026Updated 4 months ago
- One-stop on-device diffusion image-generation engine for iOS / macOS / visionOS. Embeds stable-diffusion.cpp + ggml-metal behind a tight …☆18May 19, 2026Updated 2 months ago
- Minimal Claude Code alternative powered by MLX☆47Jan 11, 2026Updated 6 months ago
- Tool for exporting Apple Neural Engine-accelerated versions of transformers models on HuggingFace Hub.☆16Jun 25, 2026Updated last month
- Triton‑style kernel toolkit for MLX plus a small upstream incubator: prototype, benchmark, and upstream fusions for Apple Silicon☆47Mar 31, 2026Updated 3 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Find out why your CoreML model isn't running on the Neural Engine!☆30Jun 18, 2024Updated 2 years ago
- Tool for visual profiling Core ML models, compatible with both package and compiled versions, including reasons for unsupported operation…☆39Jun 18, 2024Updated 2 years ago
- This repo maintains a 'cheat sheet' for LLMs that are undertrained on mlx☆33Mar 12, 2026Updated 4 months ago
- Tiny Lab is a small Apple Silicon ML research tool with a real control plane, one shipped MLX training path, and checkpoint evaluation bu…☆97Mar 10, 2026Updated 4 months ago
- ☆219Mar 24, 2026Updated 4 months ago
- MLX Implementation of Recursive Reasoning with Tiny Networks☆79Oct 11, 2025Updated 9 months ago
- Run Time Series Foundation Models on Apple Silicon☆35Feb 27, 2026Updated 5 months ago
- some sample caffemodel, prototxt, test images and pre compiled loadabes .☆14Apr 30, 2021Updated 5 years ago
- Artificial Neural Engine Machine Learning Library☆1,630Mar 10, 2026Updated 4 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Reverse engineered the Apple Neural Engine, with working Python and Objective C samples☆34Feb 1, 2026Updated 5 months ago
- Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS (E2 TTS) in MLX☆29Oct 15, 2024Updated last year
- On-device semantic search over Apple WWDC 2025 docs using MLX embeddings — SwiftUI app (WWDC OMT 2025)☆76Jun 12, 2025Updated last year
- ☆59Mar 18, 2023Updated 3 years ago
- ☆15Feb 24, 2026Updated 5 months ago
- A series of high-performance GEMM (General Matrix Multiply) implementations Iteratively optimised for H100 GPUs in Pure CUDA.☆79Feb 18, 2026Updated 5 months ago
- DeepSeek V4 Flash specific inference engine. SSD MoE expert paging (slot-bank) + disk KV cache for long agent sessions. Metal-first, narr…☆29Jul 18, 2026Updated last week
- See the device (CPU/GPU/ANE) and estimated cost for every layer in your CoreML model.☆25Oct 23, 2025Updated 9 months ago
- vibevoice real time 0.5B swift port☆31Dec 12, 2025Updated 7 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆49Jan 3, 2026Updated 6 months ago
- RayGen: Multi-Modal Dataset Reinforcement for MobileCLIP and MobileCLIP2☆40Mar 12, 2026Updated 4 months ago
- ☆18May 27, 2025Updated last year
- Flash weight streaming for MLX: run massive models larger than your RAM on Apple Silicon.☆122Jun 13, 2026Updated last month
- Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unslot…☆1,375Jun 23, 2026Updated last month
- DeepSeek 4 Flash local inference engine for Metal and CUDA with M5 optimizations.☆22May 24, 2026Updated 2 months ago
- Implementation of E2-TTS, "Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS", in MLX☆21Oct 8, 2024Updated last year