InternLM / InternLM/lmdeploy

Stuck during parallel inference.

Open
#3,057 12 comments 0 reactions 1 assignee Assigned to @lzhangzz View on GitHub
Dominant language
Python
Stars
8.1k
Forks
748
Avg merge
6d 2h
Merged PRs (30d)
54

Description

I am performing parallel inference with a batch size of 8 on a machine with 4 * A6000 GPUs. However, after running inference for a while, it gets stuck and stops responding. Meanwhile, nvidia-smi shows the following situation:

![Image](https://github.com/user-attachments/assets/326809f3-d8e8-44db-b848-7047b26f2167)

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.