alibaba / alibaba/clusterdata

How to code instance name before dataset of cluster-trace-v2018 published.

Open
#85 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
2.2k
Forks
482
PR merge metrics
No merged PRs in 30d

Description

Dear Mr. Ding,

Hello! I am a graduate student. We are very interested in the "Alibaba Cluster-trace-V2018" dataset and have downloaded the dataset.

We found that the instance name was formatted as "<_>, for example,ins_26041377" where the last four digits of the string are different and the first part of the string has several categories when we interpreted the dataset.

At the same time, we looked at some help documents on aliyun, where instance names are encoded in the format "<#><_>, for example, R2_4#281_0".

We want to know how you handle it encoding from instance A("R4_2#281_0") to instance A'("ins_26041377") before you publish this dataset. What are the mapping rules?

We are looking forward to your reply!

Sincerely Yours
Marsjiang, 20201012

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue names no files, tests, or code entry points. Start by reviewing the published Alibaba Cluster-trace-V2018 documentation and metadata for the instance identifiers, then determine whether the source-to-published-name mapping is documented. Done means providing the mapping rules or confirming that no reversible mapping is available.

Written by the indexing model from the issue text.

Assessment

Domain
data-engineering
Issue type
Documentation
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
15/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.