Mei was updated to 0.6.1, for the Ornith 1.5 model weighted avg TPS improved slightly, same for TTFT, and total task time was better. Qwen 3.6 only improved on task time and TTFT. Progress has slowed a bit, definitely harder to get gains now without making another metric worse.
github.com/tijs/mei
github.com
GitHub - tijs/mei: Native Swift/MLX OpenAI-compatible inference server for Apple Silicon
Native Swift/MLX OpenAI-compatible inference server for Apple Silicon - tijs/mei