aboutcode-org / aboutcode-org/scancode-toolkit

Accept certain known extra words in a license without losing score points

未关闭
#2,405 0 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看
license scan new feature
主要语言
Python
星标
2.6k
派生
791
平均合并
1 天 12 小时
30 天内合并 PR
5

描述

When we have a license matched exactly to a rule but there are some extra words, we do not have a 100% match score by design.
Yet there are cases where these extra words are NOT something that should degrade the score. For instance:

`The library is licensed under the GPL` as a rule
vs.
`The FOOBAR library is licensed under the GPL` as a scanned text.

Here the extra `FOOBAR` should not impact the 100% score of the rule detection.

One approach could be to maintain a set of such known words and use these in some post-processing of the license match results to keep a 100% match.

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。