How to code instance name before dataset of cluster-trace-v2018 published.
- Dominant language
- Jupyter Notebook
- Stars
- 2.2k
- Forks
- 482
- PR merge metrics
- No merged PRs in 30d
Description
Dear Mr. Ding,
Hello! I am a graduate student. We are very interested in the "Alibaba Cluster-trace-V2018" dataset and have downloaded the dataset.
We found that the instance name was formatted as "<_>, for example,ins_26041377" where the last four digits of the string are different and the first part of the string has several categories when we interpreted the dataset.
At the same time, we looked at some help documents on aliyun, where instance names are encoded in the format "<#><_>, for example, R2_4#281_0".
We want to know how you handle it encoding from instance A("R4_2#281_0") to instance A'("ins_26041377") before you publish this dataset. What are the mapping rules?
We are looking forward to your reply!
Sincerely Yours
Marsjiang, 20201012
Contributor guide
No contributing guide indexed for this repository
Research direction
The issue names no files, tests, or code entry points. Start by reviewing the published Alibaba Cluster-trace-V2018 documentation and metadata for the instance identifiers, then determine whether the source-to-published-name mapping is documented. Done means providing the mapping rules or confirming that no reversible mapping is available.
Written by the indexing model from the issue text.
Assessment
- Domain
- data-engineering
- Issue type
- Documentation
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 15/100