jakedahn / qwen3-embeddings-mlxView on GitHub
MLX-powered Qwen3 embedding server for Apple Silicon Macs. Features 0.6B/4B/8B models, 44K tokens/sec throughput, REST API, batch processing, and model hot-swapping, and more
17Aug 9, 2025Updated 11 months ago

Alternatives and similar repositories for qwen3-embeddings-mlx

Users that are interested in qwen3-embeddings-mlx are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

Are these results useful?