modelscope / modelscope/ms-swift
Support audio-flamingo-3 or not?
Open
Nobody has claimed this yet.
more models
- Dominant language
- Python
- Stars
- 15.7k
- Forks
- 1.7k
- Avg merge
- 1d 16h
- Merged PRs (30d)
- 136
Description
Describe the feature
目前可以支持nvidia模型audio-flamingo-3 的infer和微调吗?
Paste any useful information
https://huggingface.co/nvidia/audio-flamingo-3
Additional context
Add any other context or information here(其他信息可以写在这里)
Contributor guide
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
Start with the linked Hugging Face model page and inspect how ms-swift currently handles comparable audio or multimodal models. Identify the entry points and tests for inference and fine-tuning, then establish whether audio-flamingo-3 can be supported in both workflows. Done means a clear support status and documented scope for each workflow.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- huggingface, python
- Domain
- audio-video-rtc, machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Quiet
- Clarity
- Needs clarification
- Newbie friendliness
- 30/100