aymara / aymara/lima

Unaccented dictionary entries are not built

Open
#6 1 comment 0 reactions 1 assignee Claimed by @romaricb View on GitHub
bug question
Dominant language
C++
Stars
119
Forks
20
PR merge metrics
No merged PRs in 30d

Description

Currently, during resources building unaccented entries are not built (with the unaccent.pl script for example). This means that words with wrong accentuation are not recognized as they were in old LIMA versions.

Should we implement that again or just count on the orthographic correction step ?

This old method allowed to recognize strings like "un" or "UN" as instances of "U.N.".

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.