Comfy-Org / Comfy-Org/ComfyUI

Feature Request: Parallel Inference Support to Mitigate VRAM Limitations on RTX 4090

Open
#10,870 11 comments 0 reactions 0 assignees View on GitHub
Feature
Dominant language
Python
Stars
133k
Forks
15.7k
Avg merge
1d 7h
Merged PRs (30d)
158

Description

### Feature Idea

I use a single RTX 4090 for text-to-image generation. When the FPS and clarity increase, it often leads to VRAM overload. Apart from upgrading to a card with higher VRAM, such as the H100, is it possible to support parallel inference capabilities?

### Existing Solutions

_No response_

### Other

_No response_

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.