bytedance / bytedance/Protenix
Training dataset of Protein monomer predictions
- Dominant language
- Python
- Stars
- 2.1k
- Forks
- 310
- PR merge metrics
- No merged PRs in 30d
Description
I am writing to inquire about the availability of the Protein Monomer Distillation training data used in your paper. Specifically, I would like to know:
1. Is the data openly available for download? If so, could you please provide the access link?
2. If the data is not publicly available, could you kindly specify which datasets from MGnify were used in the study?
Thank you for your time and assistance. I greatly appreciate your support and look forward to your response.
Contributor guide
Research direction
Start by reviewing the paper and the project's available documentation for the Protein Monomer Distillation training data. Determine whether the data can be linked publicly or identify the MGnify datasets used; no source file or test is named, and done means providing a definitive availability answer or dataset list.
Written by the indexing model from the issue text.
Assessment
- Domain
- bioinformatics, machine-learning
- Issue type
- Documentation
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100