Pinned Loading
-
qwen38-27b-exl3-rdna4
qwen38-27b-exl3-rdna4 PublicQwen3.8-27B EXL3 on AMD gfx1201: RX 9070 XT (16 GB, tested) and Radeon AI PRO R9700 (32 GB). vLLM plugin with custom kernels: ~80 tok/s with MTP speculative decoding, int8 KV cache, 32k-64k context…
Python 2
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.