cloudflare / cloudflare/cloudflare-docs
Article suggests text-to-speech model output format is an image
Open
content:edit
documentation
product:workers-ai
September 2025
stale
- Dominant language
- MDX
- Stars
- 5.2k
- Forks
- 16.7k
- Avg merge
- 2d 6h
- Merged PRs (30d)
- 337
Description
### Existing documentation URL(s)
https://developers.cloudflare.com/workers-ai/models/aura-1/
>>>### Output
>>>The binding returns a `ReadableStream` with the image in JPEG or PNG format (check the model's output schema).
### What changes are you suggesting?
I haven't worked with this worker / model, but I'm assuming the output is audio.
### Additional information
_No response_
Contributor guide
Assessment
This issue has not been assessed yet.