abetlen / abetlen/llama-cpp-python

unknown model architecture: 'gemma-embedding'

未关闭
#2,065 5 条评论 9 个 reaction 已指派 0 人 在 GitHub 查看
主要语言
Python
星标
10.6k
派生
1.4k
PR 合并指标
PR 指标待抓取

描述

I am running llama-cpp-python Version: 0.3.16

Trying to load the recently released model [embeddinggemma-300M](https://huggingface.co/unsloth/embeddinggemma-300m-GGUF) I get the following error message:

`llama_model_load: error loading model: error loading model architecture: unknown model architecture: 'gemma-embedding'`

Support for this model architecture has been added to llama.cpp in this build: [https://github.com/ggml-org/llama.cpp/releases/tag/b6384](https://github.com/ggml-org/llama.cpp/releases/tag/b6384)

Could you please align llama-cpp-python to reflect this addition ?

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。