Skip to content
#

flash-next

Here is 1 public repository matching this topic...

The fastest way to run Qwen 3.8 Flash Next and Qwen 3.8 27B on a Mac: 125 tok/s in OpenCode on an M5 Max. Native MTP speculative decoding on Apple Silicon, exact at any temperature. OpenAI and Anthropic compatible local server.

  • Updated Sep 19, 2026
  • Python

Add this topic to your repo

To associate your repository with the flash-next topic, visit your repo's landing page and select "manage topics."

Learn more