allenai / allenai/ir_datasets

Add NTCIR Tip-of-the-Tongue

Ouverte
#306 0 commentaires 0 réactions 0 personnes assignées Voir sur GitHub
add-dataset
Langage dominant
Python
Étoiles
391
Forks
58
Métriques de merge des PR
Aucune PR mergée en 30 j

Description

**Dataset Information:*

This adds the datasets for the NTCIR 2026 Tip-of-the-Tongue shared task.

**Links to Resources:**

The shared task is described here: https://ntcir-tot.github.io/
The dataset is hosted on Zenodo: https://zenodo.org/records/18777084

**Dataset ID(s) & supported entities:**

- ntcir-tot/2026/

**Checklist**

Mark each task once completed. All should be checked prior to merging a new dataset.

- [ ] Dataset definition (in `ir_datasets/datasets/[topid].py`)
- [ ] Tests (in `tests/integration/[topid].py`)
- [ ] Metadata generated (using `ir_datasets generate_metadata` command, should appear in `ir_datasets/etc/metadata.json`)
- [ ] Documentation (in `ir_datasets/etc/[topid].yaml`)
- [ ] Documentation generated in https://github.com/seanmacavaney/ir-datasets.com/
- [ ] Downloadable content (in `ir_datasets/etc/downloads.json`)
- [ ] Download verification action (in `.github/workflows/verify_downloads.yml`). Only one needed per `topid`.
- [ ] Any small public files from NIST (or other potentially troublesome files) mirrored in https://github.com/seanmacavaney/irds-mirror/. Mirrored status properly reflected in `downloads.json`.

Guide de contribution

Aucun guide de contribution indexé pour ce dépôt

Évaluation

Cette issue n'a pas encore été évaluée.

Recevez les nouvelles issues par e-mail

Un résumé court des issues GitHub adaptées aux débutants.