Request for Official Support of HeartMuLa Music Generation Models in ComfyUI
- Dominant language
- Python
- Stars
- 133k
- Forks
- 15.7k
- Avg merge
- 1d 6h
- Merged PRs (30d)
- 155
Description
[**HeartMuLa**](https://github.com/HeartMuLa/heartlib) is a comprehensive suite of open-source music foundation models that have demonstrated impressive performance in music generation, transcription, and cross-modal retrieval. The suite includes:
1. HeartMuLa: a music language model that generates music conditioned on lyrics and tags with multilingual support including but not limited to English, Chinese, Japanese, Korean and Spanish.
2. HeartCodec: a 12.5 hz music codec with high reconstruction fidelity;
3. HeartTranscriptor: a whisper-based model specifically tuned for lyrics transcription.
4. HeartCLAP: an audio–text alignment model that establishes a unified embedding space for music descriptions and cross-modal retrieval.
Contributor guide
Research direction
No files, tests, or entry points are named. Start by reviewing the HeartMuLa, HeartCodec, HeartTranscriptor, and HeartCLAP project links and ComfyUI's model-integration conventions; the work is done when an agreed, tested form of official support for the requested suite is implemented.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- ai, machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 35/100