OpenShift 3.11中Spring Boot升级后存活探针间歇性超时求助
Spring Boot 2.6.1升级后OpenShift 3.11环境下存活探针间歇性超时问题
问题现象
- 在OpenShift 3.11环境中,将Spring Boot从1.4.5版本升级至2.6.1版本后,出现存活探针间歇性超时问题
- 对应警告信息:
Liveness probe failed: Get http://172.40.23.99:8090/monitoring/health: net/http: request canceled while waiting for connection (Client.Timeout exceeded while awaiting headers)
- 当前流量极低,内存、CPU及线程数均远低于阈值限制,问题可在不同集群计算节点复现;本次升级未修改部署配置、硬件及资源配置
当前配置信息
- 存活探针部署配置:
Liveness: http-get http://:8090/monitoring/health delay=90s timeout=3s period=50s #success=1 #failure=5
- Docker基础镜像信息:
"name": "redhat-openjdk-18/openjdk18-openshift","version": "1.12"
已做排查
- 访问日志显示健康检查请求仅耗时数毫秒(远低于存活探针设定的3秒超时时间):
- 10.131.4.1 - - [11/Sep/2022:14:22:07 +0000] "GET /monitoring/health HTTP/1.1" 200 907 13
- 10.131.4.1 - - [11/Sep/2022:14:22:57 +0000] "GET /monitoring/health HTTP/1.1" 200 907 21
- 10.131.4.1 - - [11/Sep/2022:14:23:47 +0000] "GET /monitoring/health HTTP/1.1" 200 907 9
- 10.131.4.1 - - [11/Sep/2022:14:24:37 +0000] "GET /monitoring/health HTTP/1.1" 200 907 19
- 10.131.4.1 - - [11/Sep/2022:14:25:27 +0000] "GET /monitoring/health HTTP/1.1" 200 907 8
- 尝试禁用Actuator健康检查涉及的所有组件(数据库、Redis、磁盘空间、Ping、Refresh等)后,问题仍存在
- 关键观察:扩容增加实例数,或阻断所有入站流量后,警告消失;看似问题与资源相关,存在周期性阻塞,但所有可用指标均显示正常
求助需求
恳请提供解决建议
内容的提问来源于stack exchange,提问作者Hanan Zamalin
相关产品推荐
相关产品推荐

