Distributed training
Open
- Dominant language
- C++
- Stars
- 2.9k
- Forks
- 553
- PR merge metrics
- No merged PRs in 30d
Description
I'm not clear with the given procedure for the distributed training. For the first experiment, I have partitioned the PPI dataset inti ppi_data_0.dat and ppi_data_1.dat files and loaded them to HDFS.
Can anyone give the step by step procedure to do distributed training with these files?
Contributor guide
No contributing guide indexed for this repository
Research direction
Start by reviewing the existing distributed-training instructions and the HDFS workflow for the partitioned ppi_data_0.dat and ppi_data_1.dat files. Document a step-by-step procedure for running the first experiment with those files, with completion shown by a successful distributed-training run.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- cpp, hadoop
- Domain
- data-engineering, distributed-systems, documentation, machine-learning
- Issue type
- Documentation
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100