ggml-cuda : add flash attention support for head size 88 (Llama 4 Vision) - #20375
Open
caffeinatedbits wants to merge 1 commit into
Open
caffeinatedbits wants to merge 1 commit into
caffeinatedbits wants to merge 1 commit into
The logs for this run have expired and are no longer available.
Loading