Flash Attention support for VAE
Open
Feature
- Dominant language
- Python
- Stars
- 133k
- Forks
- 15.7k
- Avg merge
- 1d 7h
- Merged PRs (30d)
- 158
Description
### Feature Idea
Can you please support flash attention in VAE (when using --use-flash-attention) instead of forcing split attention? I have tested it by modifying the files and it works fine.
### Existing Solutions
_No response_
### Other
_No response_
Contributor guide
Assessment
This issue has not been assessed yet.