anomalyco / anomalyco/opencode
webfetch returns raw XHTML markup instead of converting it
@nexxeln is already working on this.
Since Aug 28, 2026.
- Dominant language
- TypeScript
- Stars
- 209k
- Forks
- 27.5k
- PR merge metrics
- PR metrics pending
Description
Describe the bug
webfetch advertises application/xhtml+xml;q=0.9 in its own Accept header (for format: html), and isTextualMime accepts *+xml mime types, but convert() only treats text/html as HTML:
if (!contentType.includes("text/html")) return content
So a page served as application/xhtml+xml is returned as raw XML tag soup in markdown/"text" formats, with no error (because the mime check passes), leaving the model to consume low-quality markup.
Steps to reproduce
Fetch any URL serving Content-Type: application/xhtml+xml (strict XHTML documents, some CMS/doc generators) with format: "markdown". The output is the untouched <html xmlns=...><body>... source.
Affected code
packages/core/src/tool/webfetch.ts — convert() string-match only recognizes text/html.
Suggested fix
Also treat application/xhtml+xml as HTML in convert().
Environment
- opencode version: latest dev
Contributor guide
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Assessment
This issue has not been assessed yet.