deepseek-ai / deepseek-ai/Janus
In the example provided in the code, the image generation function only supports text input. Can the image generation support multimodal input? How should image input be handled?
Open
- Dominant language
- Python
- Stars
- 17.8k
- Forks
- 2.2k
- PR merge metrics
- No merged PRs in 30d
Description
In the example provided in the code, the image generation function only supports text input. Can the image generation support multimodal input? How should image input be handled?
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.