You need to enable JavaScript to run this app.
优惠活动
大模型
产品
解决方案
定价
更多

Alfresco [CONCURRENT_MAP_GET]超时错误致应用挂起求助

Alfresco集群环境频繁挂起问题排查求助

运行在Tomcat上的Alfresco应用频繁出现挂起现象,重启Alfresco服务或数据库SQL服务后可临时恢复,但数日后问题会再次复现。日志中记录大量OperationTimeoutException错误及集群缓存获取失败警告,核心错误信息如下:

13:02:19,562 ERROR [org.springframework.extensions.webscripts.AbstractRuntime] Exception from executeScript - redirecting to status template error: 10021817710 Wrapped Exception (with status template): [CONCURRENT_MAP_GET] Operation Timeout (with no response!): 20000
org.springframework.extensions.webscripts.WebScriptException: 10021817710 Wrapped Exception (with status template): [CONCURRENT_MAP_GET] Operation Timeout (with no response!): 20000
    at org.springframework.extensions.webscripts.AbstractWebScript.createStatusException(AbstractWebScript.java:1067)
    at org.springframework.extensions.webscripts.DeclarativeWebScript.execute(DeclarativeWebScript.java:171)
    ...
Caused by: com.hazelcast.core.OperationTimeoutException: [CONCURRENT_MAP_GET] Operation Timeout (with no response!): 20000
    at com.hazelcast.impl.BaseManager$ResponseQueueCall.waitAndGetResult(BaseManager.java:619)
    ...

13:02:19,656 WARN  [org.alfresco.enterprise.repo.cluster.cache.HazelcastSimpleCache] Can't get value from cluster for cache HazelcastSimpleCache[cacheName=cache.aclSharedCache], key=AclEntity[ ID=718467, version=2538900, aclId=c51ff8f7-b2de-4891-a40d-85a57e66a09e, isLatest=true, aclVersion=1, inherits=true, inheritsFrom=718466, type=2, inheritedAcl=718467, isVersioned=false, requiresVersion=false, aclChangeSet=68236281]
13:02:19,765 WARN  [org.alfresco.enterprise.repo.cluster.cache.HazelcastSimpleCache] Can't get value from cluster for cache HazelcastSimpleCache[cacheName=cache.ticketsCache], key=77b86475f7d72a489685c1b27ed4d908c9e8121c
...

问题定位与排查方向

  • 核心问题:错误根源是Alfresco依赖的Hazelcast分布式缓存节点间通信超时,导致无法获取分布式锁、ACL缓存、会话票据等关键数据,最终引发WebScript执行失败,应用挂起。
  • 可能原因及检查项:
    • 集群节点网络异常:检查节点间网络延迟、丢包率,确认防火墙/安全组未拦截Hazelcast默认通信端口(5701-5703),用ping、telnet工具验证连通性。
    • Hazelcast配置不合理:默认20秒的操作超时时间过短,或节点负载过高(CPU、内存、IO使用率超标)导致响应缓慢。可调整hazelcast.operation.timeout参数延长超时时间,同时监控节点资源使用情况。
    • 数据库性能瓶颈:数据库慢查询、连接池耗尽可能间接导致缓存加载超时,需查看数据库慢查询日志,监控连接池活跃连接数。
    • Hazelcast集群脑裂:节点间心跳中断引发集群分裂,部分节点无法同步缓存数据。查看hazelcast.log中是否有节点加入/离开的异常记录。
    • 缓存数据堆积:ACL缓存、会话票据缓存未及时清理,数据量过大导致查询超时。需调整缓存过期策略,缩短过期时间,定期清理无效缓存。

临时缓解方案

  • 修改alfresco-global.properties,延长Hazelcast操作超时时间:
    hazelcast.operation.timeout=30000
    
  • 手动清理过期会话票据缓存,避免缓存数据过度膨胀。

内容的提问来源于stack exchange,提问作者Stefan Milenkov

相关产品推荐
方舟 Agent Plan

超全模态模型 × Harness 升级,最新支持 Deepseek-V4.1-Flash、GLM-5.3 系列、Doubao-Seedream-5.0-pro、Kimi-K3 (部分), 限时 9.9 元起

最近更新时间:2026.07.07 09:22:31