ByteDance-Seed / ByteDance-Seed/FlexPrefill
Can flexprefill be applied to multimodal large models or diffusion models?
Open
- Dominant language
- Python
- Stars
- 172
- Forks
- 11
- PR merge metrics
- No merged PRs in 30d
Description
Hi, I'm curious whether the flexprefill method—designed to optimize attention prefilling in large language models—can also be adapted or applied to Diffusion models (e.g., for image or video generation)?
Are there any fundamental limitations or considerations that would prevent flexprefill from being used in these contexts?
Thank you for your great work!
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.