Hi Martin,

I believe you'll want to change auto_failback to off. Then the
resource will stay on B until it fails at which point A will take over
again.

Cheers
Guy

On 28/02/2008, Martin Fernandez <[EMAIL PROTECTED]> wrote:
> Hi All,
>
>  I have installed heartbeat 2.1.3 in 2 servers, 1 active (NODE A) and 1
>  passive (NODE B)
>  When NODE A goes down then NODE B goes up. But when NODE A came back up
>  again, NODE B fails back up to NODE A.
>
>  I want that when NODE A came back up, NODE B doesn't fail back up to it
>
>  Here my configuration files:
>
>  ha.cf
>
>  debugfile /var/log/ha-debug
>  logfile /var/log/ha-log
>  logfacility local0
>  keepalive 200ms
>  deadtime 2
>  warntime 1
>  initdead 120
>  udpport 694
>  auto_failback on
>  bcast eth0
>  node globantpbx1
>  node globantpbx2
>  crm off
>
>  haresources
>  globantpbx1 10.10.115.203 mysqld asterisk
>
>
>  Here my log file:
>
>  heartbeat[9542]: 2008/02/28_08:00:46 WARN: node globantpbx1: is dead
>  heartbeat[9542]: 2008/02/28_08:00:46 WARN: No STONITH device configured.
>  heartbeat[9542]: 2008/02/28_08:00:46 WARN: Shared disks are not protected.
>  heartbeat[9542]: 2008/02/28_08:00:46 info: Resources being acquired from
>  globantpbx1.
>  heartbeat[9542]: 2008/02/28_08:00:46 info: Link globantpbx1:eth0 dead.
>  heartbeat[9597]: 2008/02/28_08:00:46 debug: notify_world: setting SIGCHLD
>  Handler to SIG_DFL
>  harc[9597]:     2008/02/28_08:00:46 info: Running /etc/ha.d/rc.d/status
>  status
>  heartbeat[9598]: 2008/02/28_08:00:46 info: No local resources
>  [/usr/share/heartbeat/ResourceManager listkeys globantpbx2] to acquire.
>  heartbeat[9542]: 2008/02/28_08:00:46 debug: StartNextRemoteRscReq(): child
>  count 1
>  mach_down[9626]:        2008/02/28_08:00:46 info: Taking over resource group
>  10.10.115.203
>  ResourceManager[9652]:  2008/02/28_08:00:46 info: Acquiring resource group:
>  globantpbx1 10.10.115.203 mysqld asterisk
>  IPaddr[9679]:   2008/02/28_08:00:46 INFO:  Resource is stopped
>  ResourceManager[9652]:  2008/02/28_08:00:46 info: Running
>  /etc/ha.d/resource.d/IPaddr 10.10.115.203 start
>  ResourceManager[9652]:  2008/02/28_08:00:46 debug: Starting
>  /etc/ha.d/resource.d/IPaddr 10.10.115.203 start
>  IPaddr[9755]:   2008/02/28_08:00:46 INFO: Using calculated nic for
>  10.10.115.203: eth0
>  IPaddr[9755]:   2008/02/28_08:00:46 INFO: Using calculated netmask for
>  10.10.115.203: 255.255.255.0
>  IPaddr[9755]:   2008/02/28_08:00:46 DEBUG: Using calculated broadcast for
>  10.10.115.203: 10.10.115.255
>  IPaddr[9755]:   2008/02/28_08:00:46 INFO: eval ifconfig eth0:0
>  10.10.115.203netmask
>  255.255.255.0 broadcast 10.10.115.255
>  IPaddr[9755]:   2008/02/28_08:00:46 DEBUG: Sending Gratuitous Arp for
>  10.10.115.203 on eth0:0 [eth0]
>  IPaddr[9738]:   2008/02/28_08:00:46 INFO:  Success
>  INFO:  Success
>  ResourceManager[9652]:  2008/02/28_08:00:46 debug:
>  /etc/ha.d/resource.d/IPaddr 10.10.115.203 start done. RC=0
>  ResourceManager[9652]:  2008/02/28_08:00:46 info: Running
>  /etc/init.d/mysqld  start
>  ResourceManager[9652]:  2008/02/28_08:00:46 debug: Starting
>  /etc/init.d/mysqld  start
>  Starting MySQL:  [  OK  ]
>  ResourceManager[9652]:  2008/02/28_08:00:48 debug: /etc/init.d/mysqld  start
>  done. RC=0
>  mach_down[9626]:        2008/02/28_08:00:48 info:
>  /usr/share/heartbeat/mach_down: nice_failback: foreign resources acquired
>  mach_down[9626]:        2008/02/28_08:00:48 info: mach_down takeover
>  complete for node globantpbx1.
>  heartbeat[9542]: 2008/02/28_08:00:48 info: mach_down takeover complete.
>  heartbeat[9542]: 2008/02/28_08:01:19 CRIT: Cluster node globantpbx1
>  returning after partition.
>  heartbeat[9542]: 2008/02/28_08:01:19 info: For information on cluster
>  partitions, See URL: http://linux-ha.org/SplitBrain
>  heartbeat[9542]: 2008/02/28_08:01:19 WARN: Deadtime value may be too small.
>  heartbeat[9542]: 2008/02/28_08:01:19 info: See FAQ for information on tuning
>  deadtime.
>  heartbeat[9542]: 2008/02/28_08:01:19 info: URL:
>  http://linux-ha.org/FAQ#heavy_load
>  heartbeat[9542]: 2008/02/28_08:01:19 info: Link globantpbx1:eth0 up.
>  heartbeat[9542]: 2008/02/28_08:01:19 WARN: Late heartbeat: Node globantpbx1:
>  interval 34810 ms
>  heartbeat[9542]: 2008/02/28_08:01:19 info: Status update for node
>  globantpbx1: status active
>  heartbeat[10076]: 2008/02/28_08:01:19 debug: notify_world: setting SIGCHLD
>  Handler to SIG_DFL
>  harc[10076]:    2008/02/28_08:01:19 info: Running /etc/ha.d/rc.d/status
>  status
>  heartbeat[9542]: 2008/02/28_08:01:19 info: globantpbx1 wants to go standby
>  [foreign]
>  heartbeat[9542]: 2008/02/28_08:01:19 ERROR: Both machines own foreign
>  resources!
>  heartbeat[9542]: 2008/02/28_08:01:19 info: standby: acquire [foreign]
>  resources from globantpbx1
>  heartbeat[10092]: 2008/02/28_08:01:19 info: acquire local HA resources
>  (standby).
>  heartbeat[10092]: 2008/02/28_08:01:19 info: local HA resource acquisition
>  completed (standby).
>  heartbeat[9542]: 2008/02/28_08:01:19 info: Standby resource acquisition done
>  [foreign].
>  heartbeat[9542]: 2008/02/28_08:01:19 ERROR: Both machines own foreign
>  resources!
>  heartbeat[9542]: 2008/02/28_08:01:20 info: remote resource transition
>  completed.
>  heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign
>  resources!
>  heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign
>  resources!
>  heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign
>  resources!
>  heartbeat[9542]: 2008/02/28_08:01:21 info: Heartbeat shutdown in progress.
>  (9542)
>  heartbeat[10105]: 2008/02/28_08:01:21 info: Giving up all HA resources.
>  ResourceManager[10118]: 2008/02/28_08:01:21 info: Releasing resource group:
>  globantpbx1 10.10.115.203 mysqld asterisk
>  ResourceManager[10118]: 2008/02/28_08:01:21 info: Running
>  /etc/init.d/asterisk  stop
>  ResourceManager[10118]: 2008/02/28_08:01:21 debug: Starting
>  /etc/init.d/asterisk  stop
>  Shutting down asterisk: [  OK  ]
>  ResourceManager[10118]: 2008/02/28_08:01:21 debug: /etc/init.d/asterisk
>  stop done. RC=0
>  ResourceManager[10118]: 2008/02/28_08:01:21 info: Running
>  /etc/init.d/mysqld  stop
>  ResourceManager[10118]: 2008/02/28_08:01:21 debug: Starting
>  /etc/init.d/mysqld  stop
>  Stopping MySQL:  [  OK  ]
>  ResourceManager[10118]: 2008/02/28_08:01:23 debug: /etc/init.d/mysqld  stop
>  done. RC=0
>  ResourceManager[10118]: 2008/02/28_08:01:23 info: Running
>  /etc/ha.d/resource.d/IPaddr 10.10.115.203 stop
>  ResourceManager[10118]: 2008/02/28_08:01:23 debug: Starting
>  /etc/ha.d/resource.d/IPaddr 10.10.115.203 stop
>  In IP Stop
>  SIOCDELRT: No such process
>  IPaddr[10291]:  2008/02/28_08:01:23 INFO: ifconfig eth0:0 down
>  IPaddr[10274]:  2008/02/28_08:01:23 INFO:  Success
>  INFO:  Success
>  ResourceManager[10118]: 2008/02/28_08:01:23 debug:
>  /etc/ha.d/resource.d/IPaddr 10.10.115.203 stop done. RC=0
>  heartbeat[10105]: 2008/02/28_08:01:23 info: All HA resources relinquished.
>  heartbeat[9542]: 2008/02/28_08:01:25 WARN: 1 lost packet(s) for
>  [globantpbx1] [385:387]
>  heartbeat[9542]: 2008/02/28_08:01:25 info: No pkts missing from globantpbx1!
>  heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBFIFO process 9544 with
>  signal 15
>  heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBWRITE process 9545 with
>  signal 15
>  heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBREAD process 9546 with
>  signal 15
>  heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9544 exited. 3
>  remaining
>  heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9545 exited. 2
>  remaining
>  heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9546 exited. 1
>  remaining
>  heartbeat[9542]: 2008/02/28_08:01:25 info: globantpbx2 Heartbeat shutdown
>  complete.
>  heartbeat[9542]: 2008/02/28_08:01:25 info: Heartbeat restart triggered.
>  heartbeat[9542]: 2008/02/28_08:01:25 info: Restarting heartbeat.
>  heartbeat[9542]: 2008/02/28_08:01:25 info: Performing heartbeat restart
>  exec.
>  heartbeat[9542]: 2008/02/28_08:01:28 info: Version 2 support: off
>  heartbeat[9542]: 2008/02/28_08:01:28 WARN: Logging daemon is disabled
>  --enabling logging daemon is recommended
>  heartbeat[9542]: 2008/02/28_08:01:28 info: **************************
>  heartbeat[9542]: 2008/02/28_08:01:28 info: Configuration validated. Starting
>  heartbeat 2.1.3
>  heartbeat[10321]: 2008/02/28_08:01:28 info: heartbeat: version 2.1.3
>  heartbeat[10321]: 2008/02/28_08:01:28 info: Heartbeat generation: 1204107815
>  heartbeat[10321]: 2008/02/28_08:01:28 info: glib: UDP Broadcast heartbeat
>  started on port 694 (694) interface eth0
>  heartbeat[10321]: 2008/02/28_08:01:28 info: glib: UDP Broadcast heartbeat
>  closed on port 694 interface eth0 - Status: 1
>  heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_TriggerHandler: Added
>  signal manual handler
>  heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_TriggerHandler: Added
>  signal manual handler
>  heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_SignalHandler: Added
>  signal handler for signal 17
>  heartbeat[10321]: 2008/02/28_08:01:28 info: Local status now set to: 'up'
>  heartbeat[10321]: 2008/02/28_08:01:29 info: Link globantpbx1:eth0 up.
>  heartbeat[10321]: 2008/02/28_08:01:29 info: Status update for node
>  globantpbx1: status active
>  heartbeat[10321]: 2008/02/28_08:01:29 info: Link globantpbx2:eth0 up.
>
>  THANKS
>  _______________________________________________
>  Linux-HA mailing list
>  [email protected]
>  http://lists.linux-ha.org/mailman/listinfo/linux-ha
>  See also: http://linux-ha.org/ReportingProblems
>


-- 
Don't just do something...sit there!
_______________________________________________
Linux-HA mailing list
[email protected]
http://lists.linux-ha.org/mailman/listinfo/linux-ha
See also: http://linux-ha.org/ReportingProblems

Reply via email to