DetectionTeamUCAS / DetectionTeamUCAS/NAS_FPN_Tensorflow

GPU is not utilized and no progress happening

Open
#26 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
217
Forks
62
PR merge metrics
No merged PRs in 30d

Description

I have modified the code for Resnet_152_v1 and python multi_gpu_train.py. But GPU memory 1234MB only used, But GPU is not utilized ( i have set the GPU_GROUP = "0"). CPU utilization is above 230%. But it does not show the steps/epochs. I have waited for 1 hour also, no progress (refer the screenshot below) happening. Could you please suggest me what is missing in my code and config?

[
![Screen Shot 2021-01-20 at 1 24 07 AM](https://user-images.githubusercontent.com/2168986/105085810-4aca7380-5abe-11eb-98a3-9d667f15538e.png)
](url)

Contributor guide

No contributing guide indexed for this repository

Research direction

Read the modified Resnet_152_v1 code and python multi_gpu_train.py first, then inspect the GPU_GROUP="0" configuration and run the training entry point while checking whether steps or epochs advance. Done means the cause of the missing progress and GPU utilization is identified and documented with a reproducible configuration or diagnostic.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, tensorflow
Domain
computer-vision, machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.