You need to enable JavaScript to run this app.
优惠活动
大模型
产品
解决方案
定价
更多

OpenShift集群4.13.46升级至4.14.34失败:监控Operator不可用求助

OpenShift集群4.13.46升级至4.14.34失败:监控Operator不可用求助

大家好,我碰到了OpenShift集群升级的棘手问题,想请各位大佬帮忙出出主意:

我正尝试把RedHat OpenShift集群从4.13.46版本升级到4.14.34,目前33个ClusterOperators里已经完成了29个,但6个节点还一个都没开始升级,整个流程就卡住了,提示**"Cluster operator monitoring is not available"**。之前集群运行完全正常,包括之前从4.12.61升级到4.13.46的操作也顺利完成了,这次突然出了问题。

我做了一些初步诊断,结果如下:

执行oc get co monitoring的输出:

NAME         VERSION   AVAILABLE   PROGRESSING   DEGRADED   SINCE   MESSAGE
monitoring   4.13.46   False       True          True       25h     waiting for Alertmanager object changes failed: waiting for Alertmanager openshift-monitoring/main: context deadline exceeded

查看alertmanager-main-0的日志(截取关键片段):

ts=2024-08-25T13:40:42.182Z caller=dispatch.go:354 level=error component=dispatcher msg="Notify for alerts failed" num_alerts=1 err="Critical/email[0]: notify retry canceled after 10 attempts: send STARTTLS command: x509: certificate sig...

有没有朋友遇到过类似的情况?或者能给我一些排查方向的建议?

备注:内容来源于stack exchange,提问作者Nidrael

相关产品推荐
方舟 Agent Plan

超全模态模型 × Harness 升级,最新支持 Deepseek-V4.1-Flash、GLM-5.3 系列、Doubao-Seedream-5.0-pro、Kimi-K3 (部分), 限时 9.9 元起

最近更新时间:2026.04.16 08:35:32