apache / apache/cloudstack

Retrying host maintenance falsely marks already-migrated VMs as stopped

未關閉
#14,026 2 則留言 0 個 reaction 已指派 0 人 在 GitHub 檢視
component:management-server type:bug
主要語言
Java
星號
3.1k
分支
1.4k
平均合併
7 天 14 小時
30 天內合併 PR
31

描述

CLOUDSTACK VERSION
4.22.1.1, KVM hypervisor

SUMMARY
Putting a host into maintenance moves its VMs off one at a time, which
can take a while. Requesting maintenance again on the same host before
the first request finishes causes it to retry moving every VM still
listed against that host, including ones already moved. That retry
fails as expected, but for some VMs it confuses CloudStack into
thinking the VM was powered off, and it gets marked stopped even though
it was never touched and kept running the whole time.

EXPECTED
Requesting maintenance again should not retry VMs already moved.

ACTUAL
Some already-moved VMs get falsely marked stopped.

NOTE
The false "stopped" status seems to come from CloudStack losing track
of which VMs are on which host. That could plausibly happen from causes
other than this one, but we only have evidence of it via this scenario.

貢獻指南

開啟貢獻指南

研究方向

首先,在 CloudStack 4.22.1.1 中的 KVM 主機上重現重複的維護請求,然後追蹤主機維護 VM 的遷移流程,以及主機的 VM 清單如何更新。完成的標準是:第二次請求會略過已經遷移的 VM,且這些 VM 會繼續執行,而不是被標記為已停止。

由索引模型根據 Issue 內容生成。

評估

技術堆疊
java
領域
cloud, infrastructure
Issue 類型
缺陷
難度
4/5
預估耗時
3-5 天
活躍度
活躍
描述清晰度
基本清楚
新手友好度
42/100

把新 issue 寄到你的電子郵件信箱

精選適合新手參與的 GitHub issue 摘要。