Skip to content

webgpu: Fix TurboQuant quantized KV cache for batch>1 with per-batch seqlens - #29752

Merged
Jiajia Qin (qjia7) merged 6 commits into
microsoft:mainfrom
qjia7:fix/turbo-quant-batch-support
Aug 7, 2026
Merged

Jiajia Qin (qjia7) merged 6 commits into
microsoft:mainfrom
qjia7:fix/turbo-quant-batch-support

webgpu: pass TurboQuant copy length via uniform

5414ba5
Select commit
Loading
Failed to load commit list.
Sign in for the full log view

Annotations

2 warnings

The logs for this run have expired and are no longer available.