Unified provider sends `max_tokens` to OpenAI reasoning models
- Vorherrschende Sprache
- TypeScript
- Sterne
- 1.2k
- Forks
- 345
- Ø Merge
- 13 Std. 31 Min.
- Gemergte PRs (30 T.)
- 1
Beschreibung
`ai-gateway-provider`’s Unified route currently creates its models via `@ai-sdk/openai-compatible`:
```ts
const unified = createUnified();
const model = aigateway(unified('openai/o1-mini'));
```
When used with AI SDK’s standardized `maxOutputTokens`, the outgoing Chat Completions request contains:
```json
{
"model": "openai/o1-mini",
"max_tokens": 1000
}
```
OpenAI reasoning models reject this field and require `max_completion_tokens` instead:
```text
Unsupported parameter: 'max_tokens' is not supported with this model.
Use 'max_completion_tokens' instead.
```
The native Cloudflare OpenAI adapter does not have the issue:
```ts
const openai = createOpenAI();
const model = aigateway(openai.chat('o1-mini'));
```
That route uses `@ai-sdk/openai`, which performs the correct OpenAI-specific parameter mapping.
Suggested fix: have `createUnified()` supply `transformRequestBody` to its `createOpenAICompatible` instance. When Cloudflare’s model-capability metadata identifies an OpenAI reasoning model, move `max_tokens` to `max_completion_tokens`.
Beitragsleitfaden
Rechercherichtung
Start at createUnified() and its createOpenAICompatible instance, then trace how transformRequestBody receives Cloudflare’s model-capability metadata and maps standardized maxOutputTokens. Reproduce the openai/o1-mini request and verify that reasoning models send max_completion_tokens while other models retain their existing parameter behavior.
Vom Indexierungsmodell aus dem Issue-Text verfasst.
Bewertung
- Tech-Stack
- typescript
- Bereich
- ai, api
- Issue-Typ
- Bug
- Schwierigkeit
- 3/5
- Geschätzter Aufwand
- 1-2 Tage
- Aktivitätsstatus
- Aktiv
- Klarheit
- Größtenteils klar
- Anfängerfreundlichkeit
- 70/100