Skip to content

webgpu: Fix TurboQuant quantized KV cache for batch>1 with per-batch seqlens - #29752

Draft
qjia7 wants to merge 4 commits into
microsoft:mainfrom
qjia7:fix/turbo-quant-batch-support
Draft

webgpu: Fix TurboQuant quantized KV cache for batch>1 with per-batch seqlens#29752
qjia7 wants to merge 4 commits into
microsoft:mainfrom
qjia7:fix/turbo-quant-batch-support

webgpu: fix TurboQuant sequence length handling

af1c4d8
Select commit
Loading
Failed to load commit list.
Sign in for the full log view

Annotations

8 warnings
Windows GPU DML CI Pipeline
succeeded Aug 5, 2026 in 1h 44m 26s