Unaccented dictionary entries are not built
Open
bug
question
- Dominant language
- C++
- Stars
- 119
- Forks
- 20
- PR merge metrics
- No merged PRs in 30d
Description
Currently, during resources building unaccented entries are not built (with the unaccent.pl script for example). This means that words with wrong accentuation are not recognized as they were in old LIMA versions.
Should we implement that again or just count on the orthographic correction step ?
This old method allowed to recognize strings like "un" or "UN" as instances of "U.N.".
Contributor guide
Assessment
This issue has not been assessed yet.