ipython / ipython/ipyparallel

AsyncResult and px: groupby='first', groupby='merge'

未關閉
#260 0 則留言 0 個 reaction 已指派 0 人 在 GitHub 檢視
enhancement
主要語言
Jupyter Notebook
星號
2.6k
分支
1k
PR 合併指標
30 天內沒有已合併 PR

描述

The current groupby argument always displays the result of all engines.

It is useful sometimes to only display the result of a single engine -- when the execution / result is known to be symmetric. I propose adding a groupby argument:

- `first` : only display the result of the engine id 0; results from other engines are consumed and abandoned.

- `merge` : only display the unique results. For example, if engine 0 to 3 have identical results, and engine 4-8 have another identical results, then a two sets of results are displayed. The interactive debugger from intel mpi (a wrapper of gdb with result-merging) has this feature.

The merge mode is much harder in this case, because 'identical' is difficult to define -- objects have been serialized and deserialized; id is different. Using hash may be a possibility.

The `first` mode will already be very handy in trimming down the verbosity in a lot of cases.

貢獻指南

開啟貢獻指南

研究方向

Start by reading the existing AsyncResult groupby handling and the px entry point referenced in the issue. Define the behavior for groupby='first', including consuming other engine results, then inspect how existing tests cover grouped output; done means the first-engine result is shown without the other engines' output.

由索引模型根據 Issue 內容生成。

評估

技術堆疊
jupyter, python
領域
distributed-systems
Issue 類型
功能
難度
5/5
預估耗時
一週以上
活躍度
停滯
描述清晰度
基本清楚
新手友好度
25/100

把新 issue 寄到你的電子郵件信箱

精選適合新手參與的 GitHub issue 摘要。