🐛 [Bug]Torch distributed data parallel accelerate GPT2 example failing
Open
@apbose is already working on this.
Since Oct 16, 2025.
bug
Story: Multi-GPU & Distributed & Custom Ops
- Dominant language
- Python
- Stars
- 3k
- Forks
- 410
- Avg merge
- 3d 18h
- Merged PRs (30d)
- 78
Description
Torch distributed data parallel accelerate GPT2 example failing
cd examples/distributed_inference
CUDA_VISIBLE_DEVICES=0 accelerate launch data_parallel_gpt2.py <single GPU>
accelerate launch data_parallel_gpt2.py
torch 2.9.0.dev20250821+cu129
torch_tensorrt 2.10.0.dev0+0
accelerate 1.10.1
Contributor guide
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Assessment
This issue has not been assessed yet.