python / python/cpython

Improve `encodings.normalize_encoding` behaviour or docs

未關閉
#136,702 9 則留言 0 個 reaction 已指派 1 人 在 GitHub 檢視

@StanFromIreland 已經在處理了。

開始於 2025年7月16日。

stdlib topic-unicode type-bug
主要語言
Python
星號
77.2k
分支
35.9k
PR 合併指標
PR 指標待擷取

描述

Bug report

Bug description:

This is to wait for https://github.com/python/cpython/issues/55531 / https://github.com/python/cpython/pull/136643

  • This function accepts bytes since https://github.com/python/cpython/commit/98297ee7815939b124156e438b22bd652d67b5db. This is however undocumented, and untested. I propose deprecating (& later removing) it, otherwise it should be documented and tested properly.

  • This function is documented as

    encoding should be ASCII only.

    However, this is only enforced for bytes (see point 1) with a ValueError, and for strings the characters are simply removed. We should be consistent and not depend on input type. I propose enforcing this for strings too with a ValueError.

cc @malemburg (Note, absolutely no rush for this one, we should get the performance issue/C implementation sorted out first :-)

CPython versions tested on:

CPython main branch

Operating systems tested on:

No response

Linked PRs
  • gh-140030
  • gh-141345

貢獻指南

開啟貢獻指南

從這裡開始

  1. 先讀完整個 Issue,再讀專案的貢獻指南。
  2. 在 Issue 下留言說明你要接手 —— 這能避免兩個人做同樣的事。
  3. Fork 儲存庫,在一個分支上完成修改。
  4. 送出 Pull Request,並在描述裡引用這個 Issue 編號。

評估

這個 Issue 還沒有評估資料。

把新 issue 寄到你的電子郵件信箱

精選適合新手參與的 GitHub issue 摘要。