Skip to content

Pull requests: ikawrakow/ik_llama.cpp

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

NUMA Mirror implementation
#2396 opened Aug 31, 2026 by SamuelOliveirads Collaborator Loading…
Remove usless check in llama-quantize
#2394 opened Aug 31, 2026 by ikawrakow Owner Loading…
Support SWA compression with DFlash/DSpark
#2384 opened Aug 30, 2026 by SamuelOliveirads Collaborator Loading…
model: Add GLM-5.3-Flash (glm5next) runtime support
#2376 opened Aug 28, 2026 by Skelectric Contributor Loading…
2 of 4 tasks
Qwen-3.8-Next op fusions
#2375 opened Aug 28, 2026 by ikawrakow Owner Loading…
POC: QSA optimization
#2374 opened Aug 28, 2026 by ikawrakow Owner Draft
Map dense Qwen DFlash back to LLM_ARCH_DFLASH
#2370 opened Aug 28, 2026 by SamuelOliveirads Collaborator Loading…
qwen4exp: MTP (NextN) self-speculative decoding support
#2369 opened Aug 28, 2026 by jcr211 Loading…
2 of 4 tasks
Check if enough shared memory available for MLA on CUDA
#2354 opened Aug 25, 2026 by ikawrakow Owner Loading…
Dspark confidence method in spec-autotune
#2326 opened Aug 16, 2026 by SamuelOliveirads Collaborator Loading…
PR: Transfer ATSInfer Tensor Placement Solver into ik_llama.cpp
#2259 opened Aug 5, 2026 by giveen Loading…
2 of 4 tasks
ProTip! Type g p on any issue or pull request to go back to the pull request listing page.