Hi All, I have installed heartbeat 2.1.3 in 2 servers, 1 active (NODE A) and 1 passive (NODE B) When NODE A goes down then NODE B goes up. But when NODE A came back up again, NODE B fails back up to NODE A.
I want that when NODE A came back up, NODE B doesn't fail back up to it Here my configuration files: ha.cf debugfile /var/log/ha-debug logfile /var/log/ha-log logfacility local0 keepalive 200ms deadtime 2 warntime 1 initdead 120 udpport 694 auto_failback on bcast eth0 node globantpbx1 node globantpbx2 crm off haresources globantpbx1 10.10.115.203 mysqld asterisk Here my log file: heartbeat[9542]: 2008/02/28_08:00:46 WARN: node globantpbx1: is dead heartbeat[9542]: 2008/02/28_08:00:46 WARN: No STONITH device configured. heartbeat[9542]: 2008/02/28_08:00:46 WARN: Shared disks are not protected. heartbeat[9542]: 2008/02/28_08:00:46 info: Resources being acquired from globantpbx1. heartbeat[9542]: 2008/02/28_08:00:46 info: Link globantpbx1:eth0 dead. heartbeat[9597]: 2008/02/28_08:00:46 debug: notify_world: setting SIGCHLD Handler to SIG_DFL harc[9597]: 2008/02/28_08:00:46 info: Running /etc/ha.d/rc.d/status status heartbeat[9598]: 2008/02/28_08:00:46 info: No local resources [/usr/share/heartbeat/ResourceManager listkeys globantpbx2] to acquire. heartbeat[9542]: 2008/02/28_08:00:46 debug: StartNextRemoteRscReq(): child count 1 mach_down[9626]: 2008/02/28_08:00:46 info: Taking over resource group 10.10.115.203 ResourceManager[9652]: 2008/02/28_08:00:46 info: Acquiring resource group: globantpbx1 10.10.115.203 mysqld asterisk IPaddr[9679]: 2008/02/28_08:00:46 INFO: Resource is stopped ResourceManager[9652]: 2008/02/28_08:00:46 info: Running /etc/ha.d/resource.d/IPaddr 10.10.115.203 start ResourceManager[9652]: 2008/02/28_08:00:46 debug: Starting /etc/ha.d/resource.d/IPaddr 10.10.115.203 start IPaddr[9755]: 2008/02/28_08:00:46 INFO: Using calculated nic for 10.10.115.203: eth0 IPaddr[9755]: 2008/02/28_08:00:46 INFO: Using calculated netmask for 10.10.115.203: 255.255.255.0 IPaddr[9755]: 2008/02/28_08:00:46 DEBUG: Using calculated broadcast for 10.10.115.203: 10.10.115.255 IPaddr[9755]: 2008/02/28_08:00:46 INFO: eval ifconfig eth0:0 10.10.115.203netmask 255.255.255.0 broadcast 10.10.115.255 IPaddr[9755]: 2008/02/28_08:00:46 DEBUG: Sending Gratuitous Arp for 10.10.115.203 on eth0:0 [eth0] IPaddr[9738]: 2008/02/28_08:00:46 INFO: Success INFO: Success ResourceManager[9652]: 2008/02/28_08:00:46 debug: /etc/ha.d/resource.d/IPaddr 10.10.115.203 start done. RC=0 ResourceManager[9652]: 2008/02/28_08:00:46 info: Running /etc/init.d/mysqld start ResourceManager[9652]: 2008/02/28_08:00:46 debug: Starting /etc/init.d/mysqld start Starting MySQL: [ OK ] ResourceManager[9652]: 2008/02/28_08:00:48 debug: /etc/init.d/mysqld start done. RC=0 mach_down[9626]: 2008/02/28_08:00:48 info: /usr/share/heartbeat/mach_down: nice_failback: foreign resources acquired mach_down[9626]: 2008/02/28_08:00:48 info: mach_down takeover complete for node globantpbx1. heartbeat[9542]: 2008/02/28_08:00:48 info: mach_down takeover complete. heartbeat[9542]: 2008/02/28_08:01:19 CRIT: Cluster node globantpbx1 returning after partition. heartbeat[9542]: 2008/02/28_08:01:19 info: For information on cluster partitions, See URL: http://linux-ha.org/SplitBrain heartbeat[9542]: 2008/02/28_08:01:19 WARN: Deadtime value may be too small. heartbeat[9542]: 2008/02/28_08:01:19 info: See FAQ for information on tuning deadtime. heartbeat[9542]: 2008/02/28_08:01:19 info: URL: http://linux-ha.org/FAQ#heavy_load heartbeat[9542]: 2008/02/28_08:01:19 info: Link globantpbx1:eth0 up. heartbeat[9542]: 2008/02/28_08:01:19 WARN: Late heartbeat: Node globantpbx1: interval 34810 ms heartbeat[9542]: 2008/02/28_08:01:19 info: Status update for node globantpbx1: status active heartbeat[10076]: 2008/02/28_08:01:19 debug: notify_world: setting SIGCHLD Handler to SIG_DFL harc[10076]: 2008/02/28_08:01:19 info: Running /etc/ha.d/rc.d/status status heartbeat[9542]: 2008/02/28_08:01:19 info: globantpbx1 wants to go standby [foreign] heartbeat[9542]: 2008/02/28_08:01:19 ERROR: Both machines own foreign resources! heartbeat[9542]: 2008/02/28_08:01:19 info: standby: acquire [foreign] resources from globantpbx1 heartbeat[10092]: 2008/02/28_08:01:19 info: acquire local HA resources (standby). heartbeat[10092]: 2008/02/28_08:01:19 info: local HA resource acquisition completed (standby). heartbeat[9542]: 2008/02/28_08:01:19 info: Standby resource acquisition done [foreign]. heartbeat[9542]: 2008/02/28_08:01:19 ERROR: Both machines own foreign resources! heartbeat[9542]: 2008/02/28_08:01:20 info: remote resource transition completed. heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign resources! heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign resources! heartbeat[9542]: 2008/02/28_08:01:20 ERROR: Both machines own foreign resources! heartbeat[9542]: 2008/02/28_08:01:21 info: Heartbeat shutdown in progress. (9542) heartbeat[10105]: 2008/02/28_08:01:21 info: Giving up all HA resources. ResourceManager[10118]: 2008/02/28_08:01:21 info: Releasing resource group: globantpbx1 10.10.115.203 mysqld asterisk ResourceManager[10118]: 2008/02/28_08:01:21 info: Running /etc/init.d/asterisk stop ResourceManager[10118]: 2008/02/28_08:01:21 debug: Starting /etc/init.d/asterisk stop Shutting down asterisk: [ OK ] ResourceManager[10118]: 2008/02/28_08:01:21 debug: /etc/init.d/asterisk stop done. RC=0 ResourceManager[10118]: 2008/02/28_08:01:21 info: Running /etc/init.d/mysqld stop ResourceManager[10118]: 2008/02/28_08:01:21 debug: Starting /etc/init.d/mysqld stop Stopping MySQL: [ OK ] ResourceManager[10118]: 2008/02/28_08:01:23 debug: /etc/init.d/mysqld stop done. RC=0 ResourceManager[10118]: 2008/02/28_08:01:23 info: Running /etc/ha.d/resource.d/IPaddr 10.10.115.203 stop ResourceManager[10118]: 2008/02/28_08:01:23 debug: Starting /etc/ha.d/resource.d/IPaddr 10.10.115.203 stop In IP Stop SIOCDELRT: No such process IPaddr[10291]: 2008/02/28_08:01:23 INFO: ifconfig eth0:0 down IPaddr[10274]: 2008/02/28_08:01:23 INFO: Success INFO: Success ResourceManager[10118]: 2008/02/28_08:01:23 debug: /etc/ha.d/resource.d/IPaddr 10.10.115.203 stop done. RC=0 heartbeat[10105]: 2008/02/28_08:01:23 info: All HA resources relinquished. heartbeat[9542]: 2008/02/28_08:01:25 WARN: 1 lost packet(s) for [globantpbx1] [385:387] heartbeat[9542]: 2008/02/28_08:01:25 info: No pkts missing from globantpbx1! heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBFIFO process 9544 with signal 15 heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBWRITE process 9545 with signal 15 heartbeat[9542]: 2008/02/28_08:01:25 info: killing HBREAD process 9546 with signal 15 heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9544 exited. 3 remaining heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9545 exited. 2 remaining heartbeat[9542]: 2008/02/28_08:01:25 info: Core process 9546 exited. 1 remaining heartbeat[9542]: 2008/02/28_08:01:25 info: globantpbx2 Heartbeat shutdown complete. heartbeat[9542]: 2008/02/28_08:01:25 info: Heartbeat restart triggered. heartbeat[9542]: 2008/02/28_08:01:25 info: Restarting heartbeat. heartbeat[9542]: 2008/02/28_08:01:25 info: Performing heartbeat restart exec. heartbeat[9542]: 2008/02/28_08:01:28 info: Version 2 support: off heartbeat[9542]: 2008/02/28_08:01:28 WARN: Logging daemon is disabled --enabling logging daemon is recommended heartbeat[9542]: 2008/02/28_08:01:28 info: ************************** heartbeat[9542]: 2008/02/28_08:01:28 info: Configuration validated. Starting heartbeat 2.1.3 heartbeat[10321]: 2008/02/28_08:01:28 info: heartbeat: version 2.1.3 heartbeat[10321]: 2008/02/28_08:01:28 info: Heartbeat generation: 1204107815 heartbeat[10321]: 2008/02/28_08:01:28 info: glib: UDP Broadcast heartbeat started on port 694 (694) interface eth0 heartbeat[10321]: 2008/02/28_08:01:28 info: glib: UDP Broadcast heartbeat closed on port 694 interface eth0 - Status: 1 heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_TriggerHandler: Added signal manual handler heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_TriggerHandler: Added signal manual handler heartbeat[10321]: 2008/02/28_08:01:28 info: G_main_add_SignalHandler: Added signal handler for signal 17 heartbeat[10321]: 2008/02/28_08:01:28 info: Local status now set to: 'up' heartbeat[10321]: 2008/02/28_08:01:29 info: Link globantpbx1:eth0 up. heartbeat[10321]: 2008/02/28_08:01:29 info: Status update for node globantpbx1: status active heartbeat[10321]: 2008/02/28_08:01:29 info: Link globantpbx2:eth0 up. THANKS _______________________________________________ Linux-HA mailing list [email protected] http://lists.linux-ha.org/mailman/listinfo/linux-ha See also: http://linux-ha.org/ReportingProblems
