ByteDance-Seed / ByteDance-Seed/FlexPrefill

Can flexprefill be applied to multimodal large models or diffusion models?

Open
#20 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
172
Forks
11
PR merge metrics
No merged PRs in 30d

Description

Hi, I'm curious whether the flexprefill method—designed to optimize attention prefilling in large language models—can also be adapted or applied to Diffusion models (e.g., for image or video generation)?
Are there any fundamental limitations or considerations that would prevent flexprefill from being used in these contexts?

Thank you for your great work!

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.