MTPLX

MTPLX

2.24x decode TPS increase On Qwen 3.6 27B @ temp 0.6 | Native MTP Speculative Decoding On Apple Silicon With No External Drafter.

2.4k 185
Apache-2.0
last commit 2026-06-27
Website Source
Share:

About

The fastest way to run Qwen 3.8 Flash Next and Qwen 3.8 27B on a Mac: 125 tok/s in OpenCode on an M5 Max. Native MTP speculative decoding on Apple Silicon, exact at any temperature. OpenAI and Anthropic compatible local server.

MTPLX website preview

Languages

Contributors30

No features listed.

Comments Theme
slug: mtplx