elastic / elastic/roadmap

Multimodal Inference API Expansion

Open
#372 0 comments 0 reactions 1 assignee Claimed by @kapiljadhav-eis View on GitHub
Component: Elastic Inference Service product-area:search
Dominant language
No language data
Stars
6
Forks
1
PR merge metrics
No merged PRs in 30d

Description

Natively embed audio, video, and PDF content through the inference API.

**What the feature is**
An expansion of the Elastic inference API spec to support audio, video, and PDF embedding types, alongside updated OpenAI integration specs and improved default timeouts for chat and completion tasks.

**Who is this feature for**
Developers and enterprise software engineers building AI applications on non-text data.

**Expected outcome**
Handling audio, video, and PDF data today usually means juggling multiple custom tools and workflows. Native support removes that complexity, so teams can build multimodal AI applications faster without stitching together separate pipelines.

**Key user stories / use cases**
Building a search engine that finds specific moments inside video tutorials, scanning audio call recordings for compliance, and searching across large libraries of corporate PDF reports.

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.