Skip to content

Optimize Krea Vulkan - Flash Attention - #27494

Open
pwilkin wants to merge 1 commit into
ggml-org:masterfrom
pwilkin:optimize-krea-vulkan-fattn
Open

Optimize Krea Vulkan - Flash Attention#27494
pwilkin wants to merge 1 commit into
ggml-org:masterfrom
pwilkin:optimize-krea-vulkan-fattn

Conversation

@pwilkin

@pwilkin pwilkin commented Aug 21, 2026

Copy link
Copy Markdown
Member

Overview

The Flash Attention part of the Krea Vulkan optimization.

Requirements

@pwilkin
pwilkin requested review from a team, JohannesGaessler and ggerganov as code owners August 21, 2026 11:36
@github-actions github-actions Bot added testing Everything test related Vulkan Issues specific to the Vulkan backend ggml changes relating to the ggml tensor library for machine learning CUDA Related to the CUDA backend labels Aug 21, 2026
Assisted-by: OpenAI Codex
@pwilkin
pwilkin force-pushed the optimize-krea-vulkan-fattn branch from c60e95b to 4d1ec2d Compare August 21, 2026 12:31
@jeffbolznv

Copy link
Copy Markdown
Contributor

Needs a description and perf data, and some comments.

@0cc4m
0cc4m removed request for a team, JohannesGaessler and ggerganov August 22, 2026 04:27
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

CUDA Related to the CUDA backend ggml changes relating to the ggml tensor library for machine learning testing Everything test related Vulkan Issues specific to the Vulkan backend

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants