abetlen / abetlen/llama-cpp-python

Metal installation documentation

未关闭
#1,968 0 条评论 3 个 reaction 已指派 0 人 在 GitHub 查看
主要语言
Python
星标
10.6k
派生
1.4k
PR 合并指标
PR 指标待抓取

描述

I tried setting up `llama-cpp-python` in the current version `0.3.7` on my MacBook M4 Pro.
In the first step I only installed via `pip install llama-cpp-python --no-cache-dir` without specifiying the environment variable for Metal backend support.
I set the `n_gpu_layers` to `-1` to fully use the GPU.

The interesting thing is the GPU was used even without having to install the Metal backend support as stated in the current documentation. I double checked this with a fresh start and explicitely setting the `CMAKE_ARGS` env variable and did not see and difference in terms of performance or GPU usage.

This is pretty handy, because when not using `pip` for dependency management (e.g. `poetry`) passing the environment variable did not work on my side.

Maybe the documentation should be updated to state that the env arguments are no longer required? This would also reflect the documentation in https://github.com/ggml-org/llama.cpp/blob/master/docs/build.md#metal-build where it states that: "On MacOS, Metal is enabled by default"

This would also mean, that custom pre-built wheels are no longer required as well.

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。