urljoin works incorrectly for two path-relative URLs involving . and ..
未关闭
@serhiy-storchaka 已经在做这个了。
开始于 2024年11月11日。
stdlib
type-feature
- 主要语言
- Python
- 星标
- 77.2k
- 派生
- 36k
- 平均合并
- 1 天 9 小时
- 30 天内合并 PR
- 558
描述
Bug report
urllib.parse.urljoin is usually used to join a normalized absolute URL with a relative URL, and it generally works for that purpose. But if it’s used to join two path-relative URLs, it produces incorrect results in many cases when . or .. is involved.
>>> from urllib.parse import urljoin
>>> urljoin('a', 'b') # ok
'b'
>>> urljoin('a/', 'b') # ok
'a/b'
>>> urljoin('a', '.') # expected . or ./
'/'
>>> urljoin('a', '..') # expected .. or ../
'/'
>>> urljoin('..', 'b') # expected ../b
'b'
>>> urljoin('../a', 'b') # expected ../b
'b'
>>> urljoin('a', '../b') # expected ../b
'b'
>>> urljoin('../a', '../b') # expected ../../b
'b'
>>> urljoin('a/..', 'b') # expected b
'a/b'
There are also some problems when the base is a non-normalized absolute URL:
>>> urljoin('http://host/a/..', 'b') # expected http://host/b
'http://host/a/b'
Your environment
- CPython versions tested on: 3.10.5, 3.11.0rc1
- Operating system and architecture: NixOS 21.11 amd64
贡献指南
从这里开始
- 先读完整个 Issue,再读项目的贡献指南。
- 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
- Fork 仓库,在一个分支上完成修改。
- 提交 Pull Request,并在描述里引用这个 Issue 编号。
评估
这个 Issue 还没有评估数据。