Hello List,

I am maintaining a cluster where both nodes are in different
datacenters. Therefore, I don't have a serial connection for heartbeat
and must rely solely on the network connection (GB-fiber-ring)

A few days ago, we had a problem with one of the fibers so that the
spanning-tree was rebuilding. This caused an interrupt of ca 30 seconds
so that we ended up with a split-brain :(.

Can someone point me to an approach to avoid this? Is it a good idea to
play with ha.cf to have higher timeouts?

This is, what we have now:
# Thresholds (in seconds)
keepalive                       1
deadtime                        3
warntime                        2
initdead                        30

Do you recommend other technical solutions?

Best regards,
Andre

_______________________________________________
Linux-HA mailing list
[email protected]
http://lists.linux-ha.org/mailman/listinfo/linux-ha
See also: http://linux-ha.org/ReportingProblems

Reply via email to