Hi All,

I have installed heartbeat 2.1.3 in 2 servers, 1 active (NODE A) and 1
passive (NODE B)
When NODE A goes down then NODE B goes up. But when NODE A came back up
again, NODE B fails back up to NODE A.

I want that when NODE A came back up, NODE B doesn't fail back up to it

Here my configuration files:

ha.cf

debugfile /var/log/ha-debug
logfile /var/log/ha-log
logfacility local0
keepalive 200ms
deadtime 2
warntime 1
initdead 120
udpport 694
auto_failback on
bcast eth0
node globantpbx1
node globantpbx2
crm off

haresources
globantpbx1 10.10.115.203 mysqld asterisk


Here my log file:

heartbeat[9542]: 2008/02/28_08:00:46 WARN: node globantpbx1: is dead
heartbeat[9542]: 2008/02/28_08:00:46 WARN: No STONITH device configured.
heartbeat[9542]: 2008/02/28_08:00:46 WARN: Shared disks are not protected.
heartbeat[9542]: 2008/02/28_08:00:46 info: Resources being acquired from
globantpbx1.
heartbeat[9542]: 2008/02/28_08:00:46 info: Link globantpbx1:eth0 dead.
heartbeat[9597]: 2008/02/28_08:00:46 debug: notify_world: setting SIGCHLD
Handler to SIG_DFL
harc[9597]:     2008/02/28_08:00:46 info: Running /etc/ha.d/rc.d/status
status
heartbeat[9598]: 2008/02/28_08:00:46 info: No local resources
[/usr/share/heartbeat/ResourceManager listkeys globantpbx2] to acquire.
heartbeat[9542]: 2008/02/28_08:00:46 debug: StartNextRemoteRscReq(): child
count 1
mach_down[9626]:        2008/02/28_08:00:46 info: Taking over resource group
10.10.115.203
ResourceManager[9652]:  2008/02/28_08:00:46 info: Acquiring resource group:
globantpbx1 10.10.115.203 mysqld asterisk
IPaddr[9679]:   2008/02/28_08:00:46 INFO:  Resource is stopped
ResourceManager[9652]:  2008/02/28_08:00:46 info: Running
/etc/ha.d/resource.d/IPaddr 10.10.115.203 start
ResourceManager[9652]:  2008/02/28_08:00:46 debug: Starting
/etc/ha.d/resource.d/IPaddr 10.10.115.203 start
IPaddr[9755]:   2008/02/28_08:00:46 INFO: Using calculated nic for
10.10.115.203: eth0
IPaddr[9755]:   2008/02/28_08:00:46 INFO: Using calculated netmask for
10.10.115.203: 255.255.255.0
IPaddr[9755]:   2008/02/28_08:00:46 DEBUG: Using calculated broadcast for
10.10.115.203: 10.10.115.255
IPaddr[9755]:   2008/02/28_08:00:46 INFO: eval ifconfig eth0:0
10.10.115.203netmask
255.255.255.0 broadcast 10.10.115.255
IPaddr[9755]:   2008/02/28_08:00:46 DEBUG: Sending Gratuitous Arp for
10.10.115.203 on eth0:0 [eth0]
IPaddr[9738]:   2008/02/28_08:00:46 INFO:  Success
INFO:  Success
ResourceManager[9652]:  2008/02/28_08:00:46 debug:
/etc/ha.d/resource.d/IPaddr 10.10.115.203 start done. RC=0
ResourceManager[9652]:  2008/02/28_08:00:46 info: Running
/etc/init.d/mysqld  start
ResourceManager[9652]:  2008/02/28_08:00:46 debug: Starting
/etc/init.d/mysqld  start
Starting MySQL:  [  OK  ]
ResourceManager[9652]:  2008/02/28_08:00:48 debug: /etc/init.d/mysqld  start
done. RC=0
mach_down[9626]:        2008/02/28_08:00:48 info:
/usr/share/heartbeat/mach_down: nice_failback: foreign resources acquired
mach_down[9626]:        2008/02/28_08:00:48 info: mach_down takeover
complete for node globantpbx1.
heartbeat[9542]: 2008/02/28_08:00:48 info: mach_down takeover complete.
heartbeat[9542]: 2008/02/28_08:01:19 CRIT: Cluster node globantpbx1
returning after partition.
heartbeat[9542]: 2008/02/28_08:01:19 info: For information on cluster
partitions, See URL: http://linux-ha.org/SplitBrain
heartbeat[9542]: 2008/02/28_08:01:19 WARN: Deadtime value may be too small.
heartbeat[9542]: 2008/02/28_08:01:19 info: See FAQ for information on tuning
deadtime.
heartbeat[9542]: 2008/02/28_08:01:19 info: URL:
http://linux-ha.org/FAQ#heavy_load
heartbeat[9542]: 2008/02/28_08:01:19 info: Link globantpbx1:eth0 up.
heartbeat[9542]: 2008/02/28_08:01:19 WARN: Late heartbeat: Node globantpbx1:
interval 34810 ms
heartbeat[9542]: 2008/02/28_08:01:19 info: Status update for node
globantpbx1: status active
heartbeat[10076]: 2008/02/28_08:01:19 debug: notify_world: setting SIGCHLD
Handler to SIG_DFL
harc[10076]:    2008/02/28_08:01:19 info: Running /etc/ha.d/rc.d/status
status
heartbeat[9542]: 2008/02/28_08:01:19 info: globantpbx1 wants to go standby
[foreign]
heartbeat[9542]: 2008/02/28_08:01:19 ERROR: Both machines own foreign
resources!
heartbeat[9542]: 2008/02/28_08:01:19 info: standby: acquire [foreign]
resources from globantpbx1
heartbeat[10092]: 2008/02/28_08:01:19 info: acquire local HA resources
(standby).
heartbeat[10092]: 2008/02/28_08:01:19 info: local HA resource acquisition
completed (standby).
heartbeat[9542]: 2008/02/28_08:01:19 info: Standby resource acquisition done
[foreign].
heartbeat[9542]: 2008/02/28_08:01:19 ERROR: Both machines own foreign
resources!
heartbeat[9542]: 2008/02/28_08:01:20 info: remote resource transition
completed.
heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign
resources!
heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign
resources!
heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign
resources!
heartbeat[9542]: 2008/02/28_08:01:21 info: Heartbeat shutdown in progress.
(9542)
heartbeat[10105]: 2008/02/28_08:01:21 info: Giving up all HA resources.
ResourceManager[10118]: 2008/02/28_08:01:21 info: Releasing resource group:
globantpbx1 10.10.115.203 mysqld asterisk
ResourceManager[10118]: 2008/02/28_08:01:21 info: Running
/etc/init.d/asterisk  stop
ResourceManager[10118]: 2008/02/28_08:01:21 debug: Starting
/etc/init.d/asterisk  stop
Shutting down asterisk: [  OK  ]
ResourceManager[10118]: 2008/02/28_08:01:21 debug: /etc/init.d/asterisk
stop done. RC=0
ResourceManager[10118]: 2008/02/28_08:01:21 info: Running
/etc/init.d/mysqld  stop
ResourceManager[10118]: 2008/02/28_08:01:21 debug: Starting
/etc/init.d/mysqld  stop
Stopping MySQL:  [  OK  ]
ResourceManager[10118]: 2008/02/28_08:01:23 debug: /etc/init.d/mysqld  stop
done. RC=0
ResourceManager[10118]: 2008/02/28_08:01:23 info: Running
/etc/ha.d/resource.d/IPaddr 10.10.115.203 stop
ResourceManager[10118]: 2008/02/28_08:01:23 debug: Starting
/etc/ha.d/resource.d/IPaddr 10.10.115.203 stop
In IP Stop
SIOCDELRT: No such process
IPaddr[10291]:  2008/02/28_08:01:23 INFO: ifconfig eth0:0 down
IPaddr[10274]:  2008/02/28_08:01:23 INFO:  Success
INFO:  Success
ResourceManager[10118]: 2008/02/28_08:01:23 debug:
/etc/ha.d/resource.d/IPaddr 10.10.115.203 stop done. RC=0
heartbeat[10105]: 2008/02/28_08:01:23 info: All HA resources relinquished.
heartbeat[9542]: 2008/02/28_08:01:25 WARN: 1 lost packet(s) for
[globantpbx1] [385:387]
heartbeat[9542]: 2008/02/28_08:01:25 info: No pkts missing from globantpbx1!
heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBFIFO process 9544 with
signal 15
heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBWRITE process 9545 with
signal 15
heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBREAD process 9546 with
signal 15
heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9544 exited. 3
remaining
heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9545 exited. 2
remaining
heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9546 exited. 1
remaining
heartbeat[9542]: 2008/02/28_08:01:25 info: globantpbx2 Heartbeat shutdown
complete.
heartbeat[9542]: 2008/02/28_08:01:25 info: Heartbeat restart triggered.
heartbeat[9542]: 2008/02/28_08:01:25 info: Restarting heartbeat.
heartbeat[9542]: 2008/02/28_08:01:25 info: Performing heartbeat restart
exec.
heartbeat[9542]: 2008/02/28_08:01:28 info: Version 2 support: off
heartbeat[9542]: 2008/02/28_08:01:28 WARN: Logging daemon is disabled
--enabling logging daemon is recommended
heartbeat[9542]: 2008/02/28_08:01:28 info: **************************
heartbeat[9542]: 2008/02/28_08:01:28 info: Configuration validated. Starting
heartbeat 2.1.3
heartbeat[10321]: 2008/02/28_08:01:28 info: heartbeat: version 2.1.3
heartbeat[10321]: 2008/02/28_08:01:28 info: Heartbeat generation: 1204107815
heartbeat[10321]: 2008/02/28_08:01:28 info: glib: UDP Broadcast heartbeat
started on port 694 (694) interface eth0
heartbeat[10321]: 2008/02/28_08:01:28 info: glib: UDP Broadcast heartbeat
closed on port 694 interface eth0 - Status: 1
heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_TriggerHandler: Added
signal manual handler
heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_TriggerHandler: Added
signal manual handler
heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_SignalHandler: Added
signal handler for signal 17
heartbeat[10321]: 2008/02/28_08:01:28 info: Local status now set to: 'up'
heartbeat[10321]: 2008/02/28_08:01:29 info: Link globantpbx1:eth0 up.
heartbeat[10321]: 2008/02/28_08:01:29 info: Status update for node
globantpbx1: status active
heartbeat[10321]: 2008/02/28_08:01:29 info: Link globantpbx2:eth0 up.

THANKS
_______________________________________________
Linux-HA mailing list
[email protected]
http://lists.linux-ha.org/mailman/listinfo/linux-ha
See also: http://linux-ha.org/ReportingProblems

Reply via email to