Skip to content

cuda : fix flash_attn kernel to produce same results as CPU#3

Merged
FSSRepo merged 4 commits into
Pints-AI:flash-attn-cudafrom
ggml-org:flash-attn-cuda
Feb 1, 2024
Merged

cuda : fix flash_attn kernel to produce same results as CPU#3
FSSRepo merged 4 commits into
Pints-AI:flash-attn-cudafrom
ggml-org:flash-attn-cuda

cuda : increase C to 128 for better performance

ac26f27
Select commit
Loading
Failed to load commit list.

Workflow runs completed with no jobs