Azure / Azure/MachineLearningNotebooks

Pipeline ParallelRunStep: Parquet Output

未關閉
#1,810 0 則留言 1 個 reaction 已指派 0 人 在 GitHub 檢視
主要語言
Jupyter Notebook
星號
4.4k
分支
2.6k
PR 合併指標
30 天內沒有已合併 PR

描述

There seems to be a lack of documentation regarding possible output data of ParallelRunStep, as all the examples in documentation only mention writing the output as delimited file. Is there a possibility to use parquet as an alternative output format? I am currently facing the issue, that the output data contains several characters that kind of mess up AzureML's parsing ability for delimited files (newline characters within data, seperator characters within data, etc.).

---
#### Document Details

⚠ *Do not edit this section. It is required for docs.microsoft.com ➟ GitHub issue linking.*

* ID: f69044d5-213e-a764-31dd-24f8368212b7
* Version Independent ID: 23d38b1c-974a-b2fc-332a-70d7500e1751
* Content: [azureml.pipeline.steps.ParallelRunStep class - Azure Machine Learning Python](https://docs.microsoft.com/en-us/python/api/azureml-pipeline-steps/azureml.pipeline.steps.parallelrunstep?view=azure-ml-py)
* Content Source: [AzureML-Docset/stable/docs-ref-autogen/azureml-pipeline-steps/azureml.pipeline.steps.ParallelRunStep.yml](https://github.com/MicrosoftDocs/MachineLearning-Python-pr/blob/live/AzureML-Docset/stable/docs-ref-autogen/azureml-pipeline-steps/azureml.pipeline.steps.ParallelRunStep.yml)
* Service: **machine-learning**
* Sub-service: **core**
* GitHub Login: @DebFro
* Microsoft Alias: **debfro**

貢獻指南

這個儲存庫沒有索引到貢獻指南

研究方向

從 Issue 中連結的 ParallelRunStep 類別頁面及其內容來源 AzureML-Docset/stable/docs-ref-autogen/azureml-pipeline-steps/azureml.pipeline.steps.ParallelRunStep.yml 開始。檢查文件中記錄了哪些輸出格式,以及是否支援 Parquet;當文件明確回答輸出格式問題並解釋適用的行為時,即視為完成。

由索引模型根據 Issue 內容生成。

評估

技術堆疊
azure, machine-learning, python
領域
documentation, machine-learning
Issue 類型
文件
難度
3/5
預估耗時
1-2 天
活躍度
停滯
描述清晰度
基本清楚
新手友好度
35/100

把新 issue 寄到你的電子郵件信箱

精選適合新手參與的 GitHub issue 摘要。