Skip to content

opencl: quant lm_head / decode GEMV and medium-batch GEMM optimizations (speculative decoding/MTP) - #26477

Draft
wanghqc wants to merge 15 commits into
ggml-org:masterfrom
qualcomm:hq/quant-gemv-opt-intel-ready-r0730
Draft

opencl: quant lm_head / decode GEMV and medium-batch GEMM optimizations (speculative decoding/MTP)#26477
wanghqc wants to merge 15 commits into
ggml-org:masterfrom
qualcomm:hq/quant-gemv-opt-intel-ready-r0730

Commits

Commits on Aug 3, 2026