aboutcode-org / aboutcode-org/scancode-toolkit

Accept certain known extra words in a license without losing score points

Đang mở
#2,405 0 bình luận 0 reaction 0 người được giao Xem trên GitHub
license scan new feature
Ngôn ngữ chính
Python
Star
2.6k
Fork
791
Merge trung bình
1 ngày 12 giờ
Pull request đã merge (30 ngày)
5

Mô tả

When we have a license matched exactly to a rule but there are some extra words, we do not have a 100% match score by design.
Yet there are cases where these extra words are NOT something that should degrade the score. For instance:

`The library is licensed under the GPL` as a rule
vs.
`The FOOBAR library is licensed under the GPL` as a scanned text.

Here the extra `FOOBAR` should not impact the 100% score of the rule detection.

One approach could be to maintain a set of such known words and use these in some post-processing of the license match results to keep a 100% match.

Hướng dẫn đóng góp

Mở hướng dẫn đóng góp

Đánh giá

Issue này chưa được đánh giá.

Nhận issue mới trong hộp thư của bạn

Bản tóm tắt ngắn những issue GitHub phù hợp với người mới.