forked from ggml-org/llama.cpp
-
Notifications
You must be signed in to change notification settings - Fork 12
Pull requests: mxxm-t/mx-llama.cpp
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
cuda: test RCCL with a small all-reduce after startup and fall back cleanly (#7)
CUDA
ggml
#9
opened Sep 4, 2026 by
JCraigWasTaken
Loading…
speculative: sample the MTP draft and verify it by rejection sampling
documentation
Improvements or additions to documentation
server
testing
#8
opened Sep 3, 2026 by
JCraigWasTaken
Loading…
ProTip!
Follow long discussions with comments:>50.