Skip to content
#

MLX

mlx logo

MLX is a NumPy-like array framework designed for efficient and flexible machine learning on Apple silicon, brought to you by Apple machine learning research.

Here are 67 public repositories matching this topic...

The local inference server for coding agents. Pure Rust, one binary, Apple Silicon + NVIDIA. Anthropic + OpenAI APIs; the KV cache survives across turns, so turn 20 starts as fast as turn 2. OPD training on the same runtime.

  • Updated Sep 13, 2026
  • Rust

Native Rust, single-binary MLX inference server for Apple Silicon. OpenAI/Anthropic-compatible LLM serving — text, vision, audio, embeddings — with the widest weight×KV quantization matrix of any MLX server. No Python, no GGUF.

  • Updated Sep 14, 2026
  • Rust

Created by Apple

Released December 13, 2023

Followers
87 followers
Website
github.com/topics/mlx