devcontainers / devcontainers/cli

DevContainer cannot start with "HostRequirements.gpu = optional" when GPU driver installed but GPU doesn't exist.

未关闭
#319 2 条评论 0 个 reaction 已指派 1 人 已被 @chrmarti 认领 在 GitHub 查看
bug
主要语言
TypeScript
星标
3k
派生
457
平均合并
13 小时 17 分钟
30 天内合并 PR
6

描述

Thanks for working on optional GPU support. The configuration is exactly what I'm looking for!, but the current GPU check logic still not work in my use case.

According to https://github.com/devcontainers/cli/pull/173, Dev container checks if `docker info` contains nvidia runtimes for GPU driver support.

I use Linux EC2 instance with CUDA driver installed, I switch between non-GPU or GPU instance types depending on the type of current work (coding or training). So on non-GPU instance, the nvidia container runtime still available since it is installed.

A suggestion is, Nvidia container runtime also provides a cli tool `nvidia-container-cli`, can be used to get real GPU info.

on GPU instance:
```
➜ ~ nvidia-container-cli info
NVRM version: 510.47.03
CUDA version: 11.6

Device Index: 0
Device Minor: 0
Model: Tesla T4
Brand: Nvidia
GPU UUID: GPU-630ef986-c301-d90b-3581-8afa4becebc3
Bus Location: 00000000:00:1e.0
Architecture: 7.5
```

on non-GPU instance:
```
➜ ~ nvidia-container-cli info
nvidia-container-cli: initialization error: nvml error: driver not loaded
```
(exit code 1)

https://github.com/NVIDIA/nvidia-container-runtime

Thanks!

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。