aboutcode-org / aboutcode-org/scancode-toolkit

optimization for extracting email and URLs

未关闭
#595 8 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看
easy enhancement good first issue help wanted
主要语言
Python
星标
2.6k
派生
791
平均合并
1 天 12 小时
30 天内合并 PR
5

描述

In case the emails and URLs are extracted it is possible to avoid running the (expensive) regular expression and replace it with a much simpler and cheaper check to filter out large amounts of files (over 50% in my experience).

By first checking for '@' and '://' you can avoid having to run some of these checks, if these characters are not present.

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。