aboutcode-org / aboutcode-org/scancode-toolkit
optimization for extracting email and URLs
未關閉
easy
enhancement
good first issue
help wanted
- 主要語言
- Python
- 星號
- 2.6k
- 分支
- 791
- 平均合併
- 1 天 12 小時
- 30 天內合併 PR
- 5
描述
In case the emails and URLs are extracted it is possible to avoid running the (expensive) regular expression and replace it with a much simpler and cheaper check to filter out large amounts of files (over 50% in my experience).
By first checking for '@' and '://' you can avoid having to run some of these checks, if these characters are not present.
貢獻指南
評估
這個 Issue 還沒有評估資料。