mlx-apple-silicon-mlx

Installation
SKILL.md

MLX Local AI — Apple's ML Framework Powers Your Entire Fleet

Everything in this fleet runs on Apple's MLX framework. LLM inference, image generation, speech-to-text, embeddings — all MLX-native, all optimized for Apple Silicon's unified memory architecture.

The MLX stack

Capability Tool MLX usage
LLM inference Ollama MLX backend for model loading and inference on Apple Silicon
Image gen (Flux) mflux Pure MLX implementation of Flux diffusion models
Image gen (SD3) DiffusionKit MLX-native Stable Diffusion 3 and 3.5
Speech-to-text Qwen3-ASR MLX-accelerated audio transcription
Embeddings Ollama MLX backend for embedding model inference

One router. One framework. Four modalities. All local.

Setup

Installs
1
GitHub Stars
9
First Seen
Jun 18, 2026
mlx-apple-silicon-mlx — geeks-accelerator/ollama-herd