Skip to content

fix: allow mixed turbo/q8_0 KV cache types in CUDA flash-attention selection #11

fix: allow mixed turbo/q8_0 KV cache types in CUDA flash-attention selection

fix: allow mixed turbo/q8_0 KV cache types in CUDA flash-attention selection #11

Triggered via pull request July 18, 2026 22:28
Status Cancelled
Total duration 2m 10s
Artifacts

server.yml

on: pull_request
Fit to window
Zoom out
Zoom in

Annotations

5 errors
windows
Canceling since a higher priority waiting request for Server-refs/pull/3/merge-fable5/turbo-fattn-vec-routing exists
windows
The operation was canceled.
Server
Canceling since a higher priority waiting request for Server-refs/pull/3/merge-fable5/turbo-fattn-vec-routing exists
ubuntu
Canceling since a higher priority waiting request for Server-refs/pull/3/merge-fable5/turbo-fattn-vec-routing exists
ubuntu
The operation was canceled.