abetlen / abetlen/llama-cpp-python

How can I extract data from documents in JSON Output Format?

未關閉
#942 0 則留言 0 個 reaction 已指派 0 人 在 GitHub 檢視
question
主要語言
Python
星號
10.6k
分支
1.4k
PR 合併指標
PR 指標待擷取

描述

Hello,

I cannot find how can I extract data from documents in a JSON Output Format?

Just like https://hackernoon.com/unlocking-structured-json-data-with-langchain-and-gpt-a-step-by-step-tutorial for OpenAI.

I have tried to prompt through the following lines:

```
document_query = "Crear un esquema basado en este texto: " + document[0].page_content
_input = prompt.format_prompt(question=document_query)
output = chat_model(_input.to_string())
print(output)
parsed = parser.parse(output)
texts_json.append(json.dumps(parsed.dict()).encode('utf-8').decode('unicode-escape'))
```

However I cannot get just the JSON output. I get long unparsed conversations (for example, human-assistant dialogues if I use vicuna ggml-vicuna-13b-4bit.bin).

Could somebody tell me if Llamacpp can already be used for those goals?

Thanks in advance.

貢獻指南

開啟貢獻指南

評估

這個 Issue 還沒有評估資料。

把新 issue 寄到你的電子郵件信箱

精選適合新手參與的 GitHub issue 摘要。