Skip to content

project

Flash-Attention 3

Third-generation fused attention kernel implementation for transformer models, widely distributed as prebuilt binaries across hardware and framework versions.

Known aliases

  • flash-attn3

Relationships

No evidence-backed relationships are recorded.

Current clusters