Use Unicode Properties in regex normalization license expressions
未关闭
还没有人认领这个 Issue。
matching
performance
- 主要语言
- Java
- 星标
- 71
- 派生
- 44
- 平均合并
- 12 小时 54 分钟
- 30 天内合并 PR
- 7
描述
From the discussion on implementers call on 29 Oct., we could use Unicode properties in the regular expressions to simplify and possible speed up the license matching algorithms.
Reference: https://en.wikipedia.org/wiki/Unicode_character_property
贡献指南
从这里开始
- 先读完整个 Issue,再读项目的贡献指南。
- 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
- Fork 仓库,在一个分支上完成修改。
- 提交 Pull Request,并在描述里引用这个 Issue 编号。
调研方向
先定位许可证表达式规范化和匹配算法,然后查看 issue 中链接的 Unicode 字符属性参考。将现有正则表达式与提议的 Unicode 属性方法进行比较,并在修改实现之前确定预期的简化程度或性能改进。
由索引模型根据 Issue 内容生成。
评估
- 技术栈
- java
- 领域
- backend
- Issue 类型
- 重构
- 难度
- 4/5
- 预计耗时
- 3-5 天
- 活跃度
- 停滞
- 描述清晰度
- 需要澄清
- 新手友好度
- 35/100