Tencent / Tencent/WeKnora

[Question]:多模态图片信息提取问题

Open
#331 1 comment 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

question
Dominant language
Go
Stars
27.4k
Forks
3.7k
Avg merge
12h 46m
Merged PRs (30d)
277

Description

问题类别

使用问题

问题描述

多模态图片信息提取上传30余张图片,每次对话都只会筛选这5张图片参考。删除只保留1-2张相关信息的图片才能出来.

Image Image
背景信息

No response

操作系统

ubuntu 22.04

其他环境信息

No response

相关日志

已查找的资源

No response

确认事项
  • 我已经搜索了现有的 issues 和文档
  • 我已经提供了足够的信息来帮助理解问题

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

No source file, test, or entry point is named. Reproduce the conversation with more than 30 images on Ubuntu 22.04, compare the five images selected with the relevant images, and determine whether this is an intended limit or a retrieval problem. Done means the selection behavior is corrected or its limitation is clearly documented.

Written by the indexing model from the issue text.

Assessment

Tech stack
go
Domain
ai
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.