Hi Martin, I believe you'll want to change auto_failback to off. Then the resource will stay on B until it fails at which point A will take over again.
Cheers Guy On 28/02/2008, Martin Fernandez <[EMAIL PROTECTED]> wrote: > Hi All, > > I have installed heartbeat 2.1.3 in 2 servers, 1 active (NODE A) and 1 > passive (NODE B) > When NODE A goes down then NODE B goes up. But when NODE A came back up > again, NODE B fails back up to NODE A. > > I want that when NODE A came back up, NODE B doesn't fail back up to it > > Here my configuration files: > > ha.cf > > debugfile /var/log/ha-debug > logfile /var/log/ha-log > logfacility local0 > keepalive 200ms > deadtime 2 > warntime 1 > initdead 120 > udpport 694 > auto_failback on > bcast eth0 > node globantpbx1 > node globantpbx2 > crm off > > haresources > globantpbx1 10.10.115.203 mysqld asterisk > > > Here my log file: > > heartbeat[9542]: 2008/02/28_08:00:46 WARN: node globantpbx1: is dead > heartbeat[9542]: 2008/02/28_08:00:46 WARN: No STONITH device configured. > heartbeat[9542]: 2008/02/28_08:00:46 WARN: Shared disks are not protected. > heartbeat[9542]: 2008/02/28_08:00:46 info: Resources being acquired from > globantpbx1. > heartbeat[9542]: 2008/02/28_08:00:46 info: Link globantpbx1:eth0 dead. > heartbeat[9597]: 2008/02/28_08:00:46 debug: notify_world: setting SIGCHLD > Handler to SIG_DFL > harc[9597]: 2008/02/28_08:00:46 info: Running /etc/ha.d/rc.d/status > status > heartbeat[9598]: 2008/02/28_08:00:46 info: No local resources > [/usr/share/heartbeat/ResourceManager listkeys globantpbx2] to acquire. > heartbeat[9542]: 2008/02/28_08:00:46 debug: StartNextRemoteRscReq(): child > count 1 > mach_down[9626]: 2008/02/28_08:00:46 info: Taking over resource group > 10.10.115.203 > ResourceManager[9652]: 2008/02/28_08:00:46 info: Acquiring resource group: > globantpbx1 10.10.115.203 mysqld asterisk > IPaddr[9679]: 2008/02/28_08:00:46 INFO: Resource is stopped > ResourceManager[9652]: 2008/02/28_08:00:46 info: Running > /etc/ha.d/resource.d/IPaddr 10.10.115.203 start > ResourceManager[9652]: 2008/02/28_08:00:46 debug: Starting > /etc/ha.d/resource.d/IPaddr 10.10.115.203 start > IPaddr[9755]: 2008/02/28_08:00:46 INFO: Using calculated nic for > 10.10.115.203: eth0 > IPaddr[9755]: 2008/02/28_08:00:46 INFO: Using calculated netmask for > 10.10.115.203: 255.255.255.0 > IPaddr[9755]: 2008/02/28_08:00:46 DEBUG: Using calculated broadcast for > 10.10.115.203: 10.10.115.255 > IPaddr[9755]: 2008/02/28_08:00:46 INFO: eval ifconfig eth0:0 > 10.10.115.203netmask > 255.255.255.0 broadcast 10.10.115.255 > IPaddr[9755]: 2008/02/28_08:00:46 DEBUG: Sending Gratuitous Arp for > 10.10.115.203 on eth0:0 [eth0] > IPaddr[9738]: 2008/02/28_08:00:46 INFO: Success > INFO: Success > ResourceManager[9652]: 2008/02/28_08:00:46 debug: > /etc/ha.d/resource.d/IPaddr 10.10.115.203 start done. RC=0 > ResourceManager[9652]: 2008/02/28_08:00:46 info: Running > /etc/init.d/mysqld start > ResourceManager[9652]: 2008/02/28_08:00:46 debug: Starting > /etc/init.d/mysqld start > Starting MySQL: [ OK ] > ResourceManager[9652]: 2008/02/28_08:00:48 debug: /etc/init.d/mysqld start > done. RC=0 > mach_down[9626]: 2008/02/28_08:00:48 info: > /usr/share/heartbeat/mach_down: nice_failback: foreign resources acquired > mach_down[9626]: 2008/02/28_08:00:48 info: mach_down takeover > complete for node globantpbx1. > heartbeat[9542]: 2008/02/28_08:00:48 info: mach_down takeover complete. > heartbeat[9542]: 2008/02/28_08:01:19 CRIT: Cluster node globantpbx1 > returning after partition. > heartbeat[9542]: 2008/02/28_08:01:19 info: For information on cluster > partitions, See URL: http://linux-ha.org/SplitBrain > heartbeat[9542]: 2008/02/28_08:01:19 WARN: Deadtime value may be too small. > heartbeat[9542]: 2008/02/28_08:01:19 info: See FAQ for information on tuning > deadtime. > heartbeat[9542]: 2008/02/28_08:01:19 info: URL: > http://linux-ha.org/FAQ#heavy_load > heartbeat[9542]: 2008/02/28_08:01:19 info: Link globantpbx1:eth0 up. > heartbeat[9542]: 2008/02/28_08:01:19 WARN: Late heartbeat: Node globantpbx1: > interval 34810 ms > heartbeat[9542]: 2008/02/28_08:01:19 info: Status update for node > globantpbx1: status active > heartbeat[10076]: 2008/02/28_08:01:19 debug: notify_world: setting SIGCHLD > Handler to SIG_DFL > harc[10076]: 2008/02/28_08:01:19 info: Running /etc/ha.d/rc.d/status > status > heartbeat[9542]: 2008/02/28_08:01:19 info: globantpbx1 wants to go standby > [foreign] > heartbeat[9542]: 2008/02/28_08:01:19 ERROR: Both machines own foreign > resources! > heartbeat[9542]: 2008/02/28_08:01:19 info: standby: acquire [foreign] > resources from globantpbx1 > heartbeat[10092]: 2008/02/28_08:01:19 info: acquire local HA resources > (standby). > heartbeat[10092]: 2008/02/28_08:01:19 info: local HA resource acquisition > completed (standby). > heartbeat[9542]: 2008/02/28_08:01:19 info: Standby resource acquisition done > [foreign]. > heartbeat[9542]: 2008/02/28_08:01:19 ERROR: Both machines own foreign > resources! > heartbeat[9542]: 2008/02/28_08:01:20 info: remote resource transition > completed. > heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign > resources! > heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign > resources! > heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign > resources! > heartbeat[9542]: 2008/02/28_08:01:21 info: Heartbeat shutdown in progress. > (9542) > heartbeat[10105]: 2008/02/28_08:01:21 info: Giving up all HA resources. > ResourceManager[10118]: 2008/02/28_08:01:21 info: Releasing resource group: > globantpbx1 10.10.115.203 mysqld asterisk > ResourceManager[10118]: 2008/02/28_08:01:21 info: Running > /etc/init.d/asterisk stop > ResourceManager[10118]: 2008/02/28_08:01:21 debug: Starting > /etc/init.d/asterisk stop > Shutting down asterisk: [ OK ] > ResourceManager[10118]: 2008/02/28_08:01:21 debug: /etc/init.d/asterisk > stop done. RC=0 > ResourceManager[10118]: 2008/02/28_08:01:21 info: Running > /etc/init.d/mysqld stop > ResourceManager[10118]: 2008/02/28_08:01:21 debug: Starting > /etc/init.d/mysqld stop > Stopping MySQL: [ OK ] > ResourceManager[10118]: 2008/02/28_08:01:23 debug: /etc/init.d/mysqld stop > done. RC=0 > ResourceManager[10118]: 2008/02/28_08:01:23 info: Running > /etc/ha.d/resource.d/IPaddr 10.10.115.203 stop > ResourceManager[10118]: 2008/02/28_08:01:23 debug: Starting > /etc/ha.d/resource.d/IPaddr 10.10.115.203 stop > In IP Stop > SIOCDELRT: No such process > IPaddr[10291]: 2008/02/28_08:01:23 INFO: ifconfig eth0:0 down > IPaddr[10274]: 2008/02/28_08:01:23 INFO: Success > INFO: Success > ResourceManager[10118]: 2008/02/28_08:01:23 debug: > /etc/ha.d/resource.d/IPaddr 10.10.115.203 stop done. RC=0 > heartbeat[10105]: 2008/02/28_08:01:23 info: All HA resources relinquished. > heartbeat[9542]: 2008/02/28_08:01:25 WARN: 1 lost packet(s) for > [globantpbx1] [385:387] > heartbeat[9542]: 2008/02/28_08:01:25 info: No pkts missing from globantpbx1! > heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBFIFO process 9544 with > signal 15 > heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBWRITE process 9545 with > signal 15 > heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBREAD process 9546 with > signal 15 > heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9544 exited. 3 > remaining > heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9545 exited. 2 > remaining > heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9546 exited. 1 > remaining > heartbeat[9542]: 2008/02/28_08:01:25 info: globantpbx2 Heartbeat shutdown > complete. > heartbeat[9542]: 2008/02/28_08:01:25 info: Heartbeat restart triggered. > heartbeat[9542]: 2008/02/28_08:01:25 info: Restarting heartbeat. > heartbeat[9542]: 2008/02/28_08:01:25 info: Performing heartbeat restart > exec. > heartbeat[9542]: 2008/02/28_08:01:28 info: Version 2 support: off > heartbeat[9542]: 2008/02/28_08:01:28 WARN: Logging daemon is disabled > --enabling logging daemon is recommended > heartbeat[9542]: 2008/02/28_08:01:28 info: ************************** > heartbeat[9542]: 2008/02/28_08:01:28 info: Configuration validated. Starting > heartbeat 2.1.3 > heartbeat[10321]: 2008/02/28_08:01:28 info: heartbeat: version 2.1.3 > heartbeat[10321]: 2008/02/28_08:01:28 info: Heartbeat generation: 1204107815 > heartbeat[10321]: 2008/02/28_08:01:28 info: glib: UDP Broadcast heartbeat > started on port 694 (694) interface eth0 > heartbeat[10321]: 2008/02/28_08:01:28 info: glib: UDP Broadcast heartbeat > closed on port 694 interface eth0 - Status: 1 > heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_TriggerHandler: Added > signal manual handler > heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_TriggerHandler: Added > signal manual handler > heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_SignalHandler: Added > signal handler for signal 17 > heartbeat[10321]: 2008/02/28_08:01:28 info: Local status now set to: 'up' > heartbeat[10321]: 2008/02/28_08:01:29 info: Link globantpbx1:eth0 up. > heartbeat[10321]: 2008/02/28_08:01:29 info: Status update for node > globantpbx1: status active > heartbeat[10321]: 2008/02/28_08:01:29 info: Link globantpbx2:eth0 up. > > THANKS > _______________________________________________ > Linux-HA mailing list > [email protected] > http://lists.linux-ha.org/mailman/listinfo/linux-ha > See also: http://linux-ha.org/ReportingProblems > -- Don't just do something...sit there! _______________________________________________ Linux-HA mailing list [email protected] http://lists.linux-ha.org/mailman/listinfo/linux-ha See also: http://linux-ha.org/ReportingProblems
