allenai / allenai/molmo

Question Regarding Cap-F1 Evaluation Procedure

Aberta
#56 1 comentário 0 reações 0 responsáveis Ver no GitHub
Linguagem predominante
Python
Estrelas
933
Forks
96
Métricas de merge de PRs
Nenhum PR com merge em 30d

Descrição

Hi,

I am currently attempting to reproduce the Cap-F1 evaluation procedure described in the paper. I would like to ask whether it would be possible to share the prompts for computing both precision and recall, specifically for the steps involving statement matching, and consistency checking.

Thank you very much.

Guia de contribuição

Nenhum guia de contribuição indexado para este repositório

Avaliação

Esta issue ainda não foi avaliada.

Receba novas issues na sua caixa de entrada

Um resumo curto de issues do GitHub para quem está começando.