abetlen / abetlen/llama-cpp-python

igpu

未关闭
#1,709 6 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看
question
主要语言
Python
星标
10.6k
派生
1.4k
PR 合并指标
PR 指标待抓取

描述

from llama_cpp import Llama

llm = Llama(
model_path="C:\\Users\\ArabTech\\Desktop\\4\\phi-3.5-mini-instruct-q4_k_m.gguf",
n_gpu_layers=-1,
verbose=True,
)
output = llm(
"Q: Who is Napoleon Bonaparte A: ",
max_tokens=1024,
stop=["\n"] # Add a stop sequence to end generation at a newline
)
print(output)

n_gpu_layers=-1
n_gpu_layers=32

not work on igpu intel

how ofload model on igpu intel?

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。