python / python/cpython

Improve `encodings.normalize_encoding` behaviour or docs

オープン
#136,702 コメント 9 件 リアクション 0 件 担当者 1 名 GitHub で見る

@StanFromIreland がすでに取り組んでいます。

2025年7月16日 から。

stdlib topic-unicode type-bug
主要言語
Python
スター
77.2k
フォーク
35.9k
PR マージ指標
PR 指標を取得中

説明

Bug report

Bug description:

This is to wait for https://github.com/python/cpython/issues/55531 / https://github.com/python/cpython/pull/136643

  • This function accepts bytes since https://github.com/python/cpython/commit/98297ee7815939b124156e438b22bd652d67b5db. This is however undocumented, and untested. I propose deprecating (& later removing) it, otherwise it should be documented and tested properly.

  • This function is documented as

    encoding should be ASCII only.

    However, this is only enforced for bytes (see point 1) with a ValueError, and for strings the characters are simply removed. We should be consistent and not depend on input type. I propose enforcing this for strings too with a ValueError.

cc @malemburg (Note, absolutely no rush for this one, we should get the performance issue/C implementation sorted out first :-)

CPython versions tested on:

CPython main branch

Operating systems tested on:

No response

Linked PRs
  • gh-140030
  • gh-141345

コントリビューションガイド

コントリビューションガイドを開く

はじめの一歩

  1. issue を最後まで読み、次にプロジェクトのコントリビューションガイドを読みます。
  2. 着手することを issue にコメントします — 二人が同じ作業をするのを防げます。
  3. リポジトリをフォークし、ブランチを切って変更します。
  4. issue 番号を参照したプルリクエストを送ります。

評価

この issue はまだ評価されていません。

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。