Skip to content

Pull requests: ikawrakow/ik_llama.cpp

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

DSpark: gather BF16 Markov rows
#2304 opened Aug 12, 2026 by SamuelOliveirads Collaborator Loading…
Synch DFlash Tokens ID
#2303 opened Aug 12, 2026 by SamuelOliveirads Collaborator Loading…
Another minor optimization on CUDA for split mode graph
#2298 opened Aug 12, 2026 by ikawrakow Owner Loading…
CUDA: fuse rms -> add -> rms
#2297 opened Aug 12, 2026 by ikawrakow Owner Loading…
PR: Transfer ATSInfer Tensor Placement Solver into ik_llama.cpp
#2259 opened Aug 5, 2026 by giveen Loading…
2 of 4 tasks
Be able to force synchronization when copying graph inputs
#2092 opened Jul 6, 2026 by ikawrakow Owner Loading…
add --split-output-tensor / -sot CLI parameter
#2056 opened Jun 29, 2026 by Nexesenex Contributor Draft
2 of 4 tasks
implement perplexity in llama-server
#2011 opened Jun 22, 2026 by magikRUKKOLA Contributor Draft
Fix misc. expiring logit/sparam bias bugs
#1914 opened Jun 2, 2026 by dungquixote42 Contributor Loading…
2 of 4 tasks
Qwen3.5 MTP: extract selected tokens earlier
#1892 opened May 28, 2026 by ikawrakow Owner Loading…
Fix prompt cache viability
#1877 opened May 25, 2026 by zeljkokalezic Loading…
2 of 4 tasks
A GGUF MTP Extract and Merge Tool
#1849 opened May 20, 2026 by FNsi Draft
2 of 4 tasks
ProTip! no:milestone will show everything without a milestone.