ByteDance-Seed / ByteDance-Seed/m3-agent

控制过程如何加入音频与视频

Open
#14 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
1.4k
Forks
117
PR merge metrics
No merged PRs in 30d

Description

控制过程如何加入音频与视频,可以通过音频进行提问,比如问我的名字叫什么,自动人脸识别、声纹识别、ASR,从记忆图中进行搜索推理答案

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue does not name any files, tests, or entry points. First clarify the intended audio/video control flow and which of face recognition, voiceprint recognition, ASR, and memory-graph reasoning are in scope. Define acceptance criteria for each supported interaction before investigating the relevant project entry points.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
ai, audio-video-rtc, computer-vision, machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.