DetectionTeamUCAS / DetectionTeamUCAS/NAS_FPN_Tensorflow
GPU is not utilized and no progress happening
- Dominant language
- Jupyter Notebook
- Stars
- 217
- Forks
- 62
- PR merge metrics
- No merged PRs in 30d
Description
I have modified the code for Resnet_152_v1 and python multi_gpu_train.py. But GPU memory 1234MB only used, But GPU is not utilized ( i have set the GPU_GROUP = "0"). CPU utilization is above 230%. But it does not show the steps/epochs. I have waited for 1 hour also, no progress (refer the screenshot below) happening. Could you please suggest me what is missing in my code and config?
[

](url)
Contributor guide
No contributing guide indexed for this repository
Research direction
Read the modified Resnet_152_v1 code and python multi_gpu_train.py first, then inspect the GPU_GROUP="0" configuration and run the training entry point while checking whether steps or epochs advance. Done means the cause of the missing progress and GPU utilization is identified and documented with a reproducible configuration or diagnostic.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python, tensorflow
- Domain
- computer-vision, machine-learning
- Issue type
- Bug
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100