ByteDance-Seed / ByteDance-Seed/FlexPrefill

Why is the implementation in qwen different from that in llama?

Open
#14 0 comments 1 reaction 0 assignees View on GitHub
Dominant language
Python
Stars
172
Forks
11
PR merge metrics
No merged PRs in 30d

Description

Thank you for your work. But I don't understand something.Why do you need to slide past_key_value in qwen?
Image

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.