MagicStack / MagicStack/asyncpg

Connection.copy_records_to_table hanging when generator passed to records parameter raises exception with long error message

Offen
#1,354 0 Kommentare 0 Reaktionen 0 zugewiesene Personen Auf GitHub ansehen

Dieses Issue hat noch niemand übernommen.

Vorherrschende Sprache
Python
Sterne
8.1k
Forks
468
PR-Merge-Kennzahlen
Keine gemergten PRs in 30 T.

Beschreibung

When a generator passed to the records argument of Connection.copy_records_to_table raises an exception that when stringified with str(exc) produces a 9996 byte long (or greater) message it makes the COPY hang forever.

Reproduction below

#!/usr/bin/env -S uv run --script
# /// script
# requires-python = "==3.15"
# dependencies = ["asyncpg==0.31.0"]
# ///
"""Repro: asyncpg deadlocks in ROLLBACK when the records generator passed to
copy_records_to_table raises an exception whose str() exceeds PostgreSQL's
10,000-byte frontend message limit.

Run (assumes an existing PostgreSQL; only touches a session-local temp table):

    DSN=postgres://user:pass@host:5432/db uv run repro.py

Exit code 1 means the deadlock occurred.
"""

import asyncio
import os
import sys

import asyncpg

DSN = os.environ.get("DSN", "postgres://postgres:postgres@localhost:5432/postgres")
HANG_TIMEOUT = float(os.environ.get("HANG_TIMEOUT", "15"))

# Length in bytes of the error message raised by the record generator.
# > 9995 bytes => PostgreSQL rejects the CopyFail frame ("invalid message
# length") and the deadlock triggers; at or below, the run fails cleanly.
# The threshold is exact: the frame's int32 length field (4 bytes, counted in
# itself) + payload + NUL terminator must not exceed PQ_SMALL_MESSAGE_LIMIT
# (10,000), so 10000 - 4 - 1 = 9995.
ERROR_BYTES = int(os.environ.get("ERROR_BYTES", "9996"))


def records():
    for i in range(100):
        yield (i,)
    raise ValueError("A" * ERROR_BYTES)


async def run_case():
    conn = await asyncpg.connect(DSN)
    try:
        async with conn.transaction():
            await conn.execute("CREATE TEMP TABLE repro_t (b int)")
            await conn.copy_records_to_table("repro_t", records=records())
    except ValueError:
        # clean failure, bug not reproduced
        pass
    finally:
        await conn.close()


async def main():
    task = asyncio.create_task(run_case())
    _, pending = await asyncio.wait([task], timeout=HANG_TIMEOUT)
    if pending:
        print("Statement timed out, bug reproduced", file=sys.stderr)
        return 1
    print("Statement executed before timeout, bug not reproduced", file=sys.stderr)
    return 0


if __name__ == "__main__":
    os._exit(asyncio.run(main()))

Beitragsleitfaden

Für dieses Repository ist kein Beitragsleitfaden indexiert

Erste Schritte

  1. Lies das ganze Issue und danach den Beitragsleitfaden des Projekts.
  2. Schreib ins Issue, dass du es übernimmst — das erspart doppelte Arbeit.
  3. Forke das Repository und arbeite in einem Branch.
  4. Öffne einen Pull Request, der die Issue-Nummer nennt.

Rechercherichtung

Beginne bei Connection.copy_records_to_table und führe repro.py gegen PostgreSQL aus, wobei du ERROR_BYTES um die Schwelle von 9995 Byte variierst. Verfolge den Pfad der Generator-Exception und des COPY-Rollbacks; abgeschlossen ist die Aufgabe, wenn die Exception sauber beendet wird, ohne zu hängen, auch bei Nachrichten an oder über der Schwelle.

Vom Indexierungsmodell aus dem Issue-Text verfasst.

Bewertung

Tech-Stack
postgresql, python
Bereich
databases
Issue-Typ
Bug
Schwierigkeit
4/5
Geschätzter Aufwand
3-5 Tage
Aktivitätsstatus
Aktiv
Klarheit
Größtenteils klar
Anfängerfreundlichkeit
48/100

Neue Issues direkt in Ihr Postfach

Eine kurze Übersicht über anfängerfreundliche GitHub-Issues.