abetlen / abetlen/llama-cpp-python
Multi-arch support for pre-built cpu wheel
未關閉
enhancement
- 主要語言
- Python
- 星號
- 10.6k
- 分支
- 1.4k
- PR 合併指標
- PR 指標待擷取
描述
Currently pre-built cpu wheels are compiled for the lowest common denominator architecture. If we can compile multiple versions of the llama.cpp library with different accelerations compiled in each we can bundle all of them into the pre-built wheels and dynamically choose one based on the host cpu.
貢獻指南
評估
這個 Issue 還沒有評估資料。