speech encoder integration
Nobody has claimed this yet.
- Dominant language
- C++
- Stars
- 40.3k
- Forks
- 3.7k
- PR merge metrics
- No merged PRs in 30d
Description
Hi,
I would like to know, the possibilities for integrating the speech encoder to turn this model input Speech + Text input instead of Text only input. Any information regarding the integration of speech encoder or trainining with speech instructions would really helpful a lot.
Thanks & Regards,
Pavan.
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
No file, test, or entry point is named. Start by reviewing BitNet's current text-only model input path and determine what speech encoder and speech-instruction training requirements would need to be specified; the issue is complete only when a concrete integration scope and implementation plan are documented.
Written by the indexing model from the issue text.
Assessment
- Domain
- machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 15/100