urllib.parse.parse_qs should not parse a query string containing illegal characters
未关闭
还没有人认领这个 Issue。
stdlib
type-bug
- 主要语言
- Python
- 星标
- 77.2k
- 派生
- 35.9k
- PR 合并指标
- PR 指标待抓取
描述
Bug report
Bug description:
urllib.parse.parse_qsl will happily parse a query string containing '#'.
from urllib.parse import parse_qsl
parse_qsl('foo=#', strict_parsing=True)
Output is [('foo', '#')] .
But 'foo=#' is an invalid query string according to RFC 3986. Similarly, '[', and ']' are excluded from the set of valid query characters, but parse_qsl parses strings like 'foo=[' and 'foo=]' . In the absence of any allocation of responsibility, this looks like a bug to me.
CPython versions tested on:
3.13
Operating systems tested on:
macOS
Linked PRs
- gh-152530
贡献指南
从这里开始
- 先读完整个 Issue,再读项目的贡献指南。
- 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
- Fork 仓库,在一个分支上完成修改。
- 提交 Pull Request,并在描述里引用这个 Issue 编号。
调研方向
从 urllib.parse.parse_qsl 开始,检查查询字符如何根据 RFC 3986 进行验证。先查看现有的解析测试和链接的 PR gh-152530;当诸如 '#'、'[' 和 ']' 之类的非法字符按照已达成一致的行为得到一致处理,并由测试覆盖时,即视为完成。
由索引模型根据 Issue 内容生成。
评估
- 技术栈
- python
- 领域
- backend-api-design
- Issue 类型
- 缺陷
- 难度
- 3/5
- 预计耗时
- 1-2 天
- 活跃度
- 停滞
- 描述清晰度
- 描述清楚
- 新手友好度
- 35/100