alibaba / alibaba/euler

Distributed training

Open
#179 4 comments 0 reactions 0 assignees View on GitHub
Dominant language
C++
Stars
2.9k
Forks
553
PR merge metrics
No merged PRs in 30d

Description

I'm not clear with the given procedure for the distributed training. For the first experiment, I have partitioned the PPI dataset inti ppi_data_0.dat and ppi_data_1.dat files and loaded them to HDFS.

Can anyone give the step by step procedure to do distributed training with these files?

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by reviewing the existing distributed-training instructions and the HDFS workflow for the partitioned ppi_data_0.dat and ppi_data_1.dat files. Document a step-by-step procedure for running the first experiment with those files, with completion shown by a successful distributed-training run.

Written by the indexing model from the issue text.

Assessment

Tech stack
cpp, hadoop
Domain
data-engineering, distributed-systems, documentation, machine-learning
Issue type
Documentation
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.