Skip to content

Pull requests: GenerelSchwerz/llama.cpp

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

docs : link CLAUDE.md to AGENTS.md documentation Improvements or additions to documentation
#74 opened Sep 6, 2026 by GenerelSchwerz Owner Loading…
llama-bench: add kv-offload/workspace sweep flags documentation Improvements or additions to documentation examples
#70 opened Sep 4, 2026 by Piggidragon Draft
4 tasks done
llama : add an attention split separate from the tensor split devops documentation Improvements or additions to documentation ggml testing
#69 opened Sep 3, 2026 by Piggidragon Loading…
ggml-meta : split a host-resident KV cache by head devops documentation Improvements or additions to documentation ggml testing
#66 opened Sep 3, 2026 by Piggidragon Loading…
ggml-cuda : reuse routing IDs across MoE siblings
#42 opened Aug 26, 2026 by GenerelSchwerz Owner Draft
1 task done
sched: pipeline the delivery of a host-resident KV cache documentation Improvements or additions to documentation examples ggml testing
#39 opened Aug 26, 2026 by Piggidragon Loading…
2 tasks
kv: replace eligible dense causal masks with compact prefixes CUDA documentation Improvements or additions to documentation ggml server testing
#7 opened Aug 21, 2026 by GenerelSchwerz Owner Loading…
ProTip! Type g p on any issue or pull request to go back to the pull request listing page.