AOSSIE-Org / AOSSIE-Org/Ell-ena

summarize-transcription Edge Function has no input size limit - oversized transcriptions exhaust function memory and Supabase billing

オープン
#309 コメント 1 件 リアクション 0 件 担当者 0 名 GitHub で見る
主要言語
Dart
スター
54
フォーク
110
PR マージ指標
30日以内にマージされた PR はありません

説明

## Problem

The `supabase/functions/summarize-transcription/` Edge Function accepts
meeting transcription text and passes it to the Gemini API for summarization.
There is no maximum input size check before the function processes the text.

Supabase Edge Functions run on Deno with a memory limit (typically 150MB).
A very long meeting transcription (e.g. a 4-hour all-hands meeting producing
100,000+ tokens) can:

1. Exceed the Gemini API's context window, causing the API call to fail with
an unclear error.
2. Exhaust the Edge Function's memory limit, causing the function to crash
with a 500 error and no useful error message returned to the client.
3. Generate outsized Gemini API charges if the input is not truncated before
billing begins.

## Suggested Fix

1. Add a hard character limit (e.g. 500,000 characters) to the transcription
input before processing:
```typescript
const MAX_INPUT_CHARS = 500_000;
if (transcription.length > MAX_INPUT_CHARS) {
return new Response(JSON.stringify({ error: "Transcription too long" }), { status: 413 });
}
```
2. For long transcriptions, implement chunked summarization: split the text
into overlapping chunks, summarize each, then summarize the summaries.
3. Return HTTP 413 (Payload Too Large) with a clear message when the limit
is exceeded.
4. Document the maximum supported transcription length in `BACKEND.md`.

コントリビューションガイド

このリポジトリのコントリビューションガイドは索引されていません

評価

この issue はまだ評価されていません。

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。