Skip to content

fix: allow mixed turbo/q8_0 KV cache types in CUDA flash-attention selection #12

fix: allow mixed turbo/q8_0 KV cache types in CUDA flash-attention selection

fix: allow mixed turbo/q8_0 KV cache types in CUDA flash-attention selection #12

Triggered via pull request July 18, 2026 22:30
Status Success
Total duration 20m 38s
Artifacts

server.yml

on: pull_request
Fit to window
Zoom out
Zoom in