documentation for indexing trec collections [LUCENE-2927]
- Dominant language
- Java
- Stars
- 3.6k
- Forks
- 1.4k
- Avg merge
- 2d 11h
- Merged PRs (30d)
- 88
Description
In #2614 there are great improvements for parsing the various format differences
of some common TREC collections.
It would be nice to engage new users (especially students/academics/experimenters) if
we had some simple instructions on how to do a TREC run on the lucene website.
For reference see https://issues.apache.org/jira/browse/LUCENE-1540?focusedCommentId=12995865&page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel#comment-12995865
---
Migrated from [LUCENE-2927](https://issues.apache.org/jira/browse/LUCENE-2927) by Robert Muir (@rmuir)
Contributor guide
Research direction
Start with the Lucene website and review the parsing improvements referenced in #2614, then read the example discussion in LUCENE-1540. Done means the website contains simple, user-facing instructions for running a TREC collection indexing and search run.
Written by the indexing model from the issue text.
Assessment
- Domain
- documentation, search
- Issue type
- Documentation
- Difficulty
- 3/5
- Estimated time
- 1-2 days
- Activity status
- Stale
- Clarity
- Mostly clear
- Newbie friendliness
- 42/100