Skip to content

Pull requests: unslothai/llama.cpp

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

ggml-cuda: use cudaMemcpyDefault in the ggml_cuda_cpy 2D fast path bug Something isn't working
#157 opened Aug 31, 2026 by danielhanchen Member Loading…
MTP for Qwen3.8-Flash-Next
#144 opened Aug 30, 2026 by danielhanchen Member Loading…
llama: batched readahead for lazily read gather tables
#137 opened Aug 28, 2026 by danielhanchen Member Loading…
ProTip! Follow long discussions with comments:>50.