deepseek-ai / deepseek-ai/DeepSeek-V2

Could we have scores for `LongBookQA Eng` and `LongBookSum Eng`

Open
#4 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
No language data
Stars
5k
Forks
550
PR merge metrics
No merged PRs in 30d

Description

Some results pasted below from this [link](https://github.com/OpenBMB/InfiniteBench/tree/main):

| Task Name | GPT-4 | YaRN-Mistral-7B | Kimi-Chat | Claude 2 | Yi-6B-200K | Yi-34B-200K | Chatglm3-6B-128K |
| ---------------- | ------ | --------------- | --------- | -------- | -----------| -----------| -----------|
| En.Sum | 14.73% | 9.09% | 17.93% | 14.45% | < 5% | < 5% |< 5% |
| En.QA | 22.22% | 9.55% | 16.52% | 11.97% | 9.20% | 12.17% |< 5% |

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.