forked from ggml-org/llama.cpp
-
Notifications
You must be signed in to change notification settings - Fork 13
Pull requests: aifoundry-org/llama.cpp
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
et-backend: F32 vecdot GEMV + matrix-engine GEMM
Apple Metal
build
conversion
CUDA
devops
documentation
Improvements or additions to documentation
ggml
Hexagon
model
mtmd
OpenCL
server/ui
server
SYCL
testing
vendor
Vulkan
WebGPU
#29
opened Jul 23, 2026 by
RehanQasim-dev
•
Draft
et-backend: F16 vecdot GEMV + matrix-engine GEMM
Apple Metal
build
conversion
CUDA
devops
documentation
Improvements or additions to documentation
ggml
Hexagon
model
mtmd
OpenCL
server/ui
server
SYCL
testing
vendor
Vulkan
WebGPU
#28
opened Jul 23, 2026 by
RehanQasim-dev
•
Draft
feat: add SpatialLMQwenForCausalLM converter
conversion
#22
opened Jul 21, 2026 by
CodeDoes
Loading…
fix: stable fgb.ps dot product and N-split for 2048 harts
ggml
#21
opened Jul 21, 2026 by
CodeDoes
Loading…
ET backend: N-split infrastructure for Q4_0, Q4_K, Q8_0 MUL_MAT kernels
ggml
#20
opened Jul 21, 2026 by
CodeDoes
Loading…
ProTip!
Type g p on any issue or pull request to go back to the pull request listing page.