Skip to content

ggml-cuda: enable concurrent streams for linear attention - #21897

Open
am17an wants to merge 3 commits into
ggml-org:masterfrom
am17an:linear-attn-conc
Open

ggml-cuda: enable concurrent streams for linear attention#21897
am17an wants to merge 3 commits into
ggml-org:masterfrom
am17an:linear-attn-conc

use alloc_deps

56339ee
Select commit
Loading
Failed to load commit list.
Sign in for the full log view