Azure / Azure/MachineLearningNotebooks

Not recognizing libcudart when running commands

Open
#1,634 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
4.4k
Forks
2.6k
PR merge metrics
No merged PRs in 30d

Description

I have a GPU compute (Standard_NC6s_v3), and I want to train a model using that. The kernel/environment (Azure ML 3.8) has tensorflow 2.3 install by default, and CUDA 11. But when I run a few commands (example: python3 -c "import tensorflow as tf;print(tf.__version__)" to get the TF version), it gives me a warning:

```
2021-11-16 13:23:10.274930: W tensorflow/stream_executor/platform/default/dso_loader.cc:59] Could not load dynamic library 'libcudart.so.10.1'; dlerror: libcudart.so.10.1: cannot open shared object file: No such file or directory
2021-11-16 13:23:10.274970: I tensorflow/stream_executor/cuda/cudart_stub.cc:29] Ignore above cudart dlerror if you do not have a GPU set up on your machine.

```

Why is it looking for libcudart.so.10.1 when version 11 is installed? I also can't manually install CUDA 10.1. I tried resetting the LD_LIBRARY_PATH variable, using `export LD_LIBRARY_PATH=/usr/local/cuda-11.1/targets/x86_64-linux/lib` but that didn't work either.

Checking `ls /usr/local/ `using the terminal, I have three cuda folders; cuda, cuda 11, and cuda 11.1.

How can I bypass that error/warning?

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.