AsyncResult and px: groupby='first', groupby='merge'
- 主要語言
- Jupyter Notebook
- 星號
- 2.6k
- 分支
- 1k
- PR 合併指標
- 30 天內沒有已合併 PR
描述
The current groupby argument always displays the result of all engines.
It is useful sometimes to only display the result of a single engine -- when the execution / result is known to be symmetric. I propose adding a groupby argument:
- `first` : only display the result of the engine id 0; results from other engines are consumed and abandoned.
- `merge` : only display the unique results. For example, if engine 0 to 3 have identical results, and engine 4-8 have another identical results, then a two sets of results are displayed. The interactive debugger from intel mpi (a wrapper of gdb with result-merging) has this feature.
The merge mode is much harder in this case, because 'identical' is difficult to define -- objects have been serialized and deserialized; id is different. Using hash may be a possibility.
The `first` mode will already be very handy in trimming down the verbosity in a lot of cases.
貢獻指南
研究方向
Start by reading the existing AsyncResult groupby handling and the px entry point referenced in the issue. Define the behavior for groupby='first', including consuming other engine results, then inspect how existing tests cover grouped output; done means the first-engine result is shown without the other engines' output.
由索引模型根據 Issue 內容生成。
評估
- 技術堆疊
- jupyter, python
- 領域
- distributed-systems
- Issue 類型
- 功能
- 難度
- 5/5
- 預估耗時
- 一週以上
- 活躍度
- 停滯
- 描述清晰度
- 基本清楚
- 新手友好度
- 25/100