abetlen / abetlen/llama-cpp-python

How can I extract data from documents in JSON Output Format?

Offen
#942 0 Kommentare 0 Reaktionen 0 zugewiesene Personen Auf GitHub ansehen
question
Vorherrschende Sprache
Python
Sterne
10.6k
Forks
1.4k
PR-Merge-Kennzahlen
PR-Kennzahlen ausstehend

Beschreibung

Hello,

I cannot find how can I extract data from documents in a JSON Output Format?

Just like https://hackernoon.com/unlocking-structured-json-data-with-langchain-and-gpt-a-step-by-step-tutorial for OpenAI.

I have tried to prompt through the following lines:

```
document_query = "Crear un esquema basado en este texto: " + document[0].page_content
_input = prompt.format_prompt(question=document_query)
output = chat_model(_input.to_string())
print(output)
parsed = parser.parse(output)
texts_json.append(json.dumps(parsed.dict()).encode('utf-8').decode('unicode-escape'))
```

However I cannot get just the JSON output. I get long unparsed conversations (for example, human-assistant dialogues if I use vicuna ggml-vicuna-13b-4bit.bin).

Could somebody tell me if Llamacpp can already be used for those goals?

Thanks in advance.

Beitragsleitfaden

Beitragsleitfaden öffnen

Bewertung

Dieses Issue wurde noch nicht bewertet.

Neue Issues direkt in Ihr Postfach

Eine kurze Übersicht über anfängerfreundliche GitHub-Issues.