On Tue, Aug 28, 2007 at 11:47:23AM -0500, Bret Baptist wrote:
> On Saturday 25 August 2007 10:48:46 pm you wrote:
> > I have a new HA cluster of two machines. After fiddling with the
> > configuration for a while, I was able to get two resources configured:
> > IPaddr and ldirectord. That functions properly; I can use crm_mon and
> > crm_standby to watch it function. For further configuration, I would
> > really like to use the hb_gui, but every time I use it, the
> > application stops responding. I can get as far as logging in, querying
> > the crm/cib. I can browse all the current attributes, properties, and
> > so forth. However, any time I want to make a change, such as add
> > another resource, it locks up.
> >
> > I've tried running hb_gui locally on the active node which is DC at
> > the time, and over a SSH-tunneled X session. I don't see anything
> > regarding the error in the log files, and there aren't any messages
> > sent to STDERR by hb_gui itself.
> >
> > Whenever this happens, I can't get the heartbeat service to restart.
> > Something (I believe it was mgmtd last time) won't quit without a
> > `kill -9`. I find the easiest solution to force a reboot and let it
> > start over. When it comes back up, it functions properly again.
> >
> > Should I not bother with the hb_gui, and use the crm_* tools instead?
> > I'm not sure if that's practical currently.
> >
> > Thanks.
> 
> 
> I am having these exact same issues.  I am using 2.1.2 compiled for ubuntu 
> 7.04 server.
> 
> When I try to add a resource using the hb_gui I get these messages in syslog:
> Aug 28 11:42:38 hope lrmd: [17831]: debug: stonithRA plugin: provider 
> attribute is not needed and will be ignored.
> Aug 28 11:42:38 hope lrmd: [17831]: WARN: stonithRA plugin: cannot get 
> shortdesc segment of apcmaster's metadata.
> Aug 28 11:42:38 hope lrmd: [17831]: debug: stonithRA plugin: provider 
> attribute is not needed and will be ignored.
> Aug 28 11:42:38 hope lrmd: [17831]: WARN: stonithRA plugin: cannot get 
> shortdesc segment of apcmastersnmp's metadata.
> Aug 28 11:42:38 hope lrmd: [17831]: debug: stonithRA plugin: provider 
> attribute is not needed and will be ignored.
> Aug 28 11:42:38 hope lrmd: [17831]: WARN: stonithRA plugin: cannot get 
> shortdesc segment of apcsmart's metadata.
> Aug 28 11:42:38 hope lrmd: [17831]: debug: stonithRA plugin: provider 
> attribute is not needed and will be ignored.
> Aug 28 11:42:38 hope lrmd: [17831]: WARN: stonithRA plugin: cannot get 
> shortdesc segment of baytech's metadata.
> Aug 28 11:42:38 hope lrmd: [17831]: debug: stonithRA plugin: provider 
> attribute is not needed and will be ignored.
> Aug 28 11:42:38 hope lrmd: [17831]: ERROR: cl_free: Bad magic number in 
> object 
> at 0x805e350
> Aug 28 11:42:38 hope lrmd: [17831]: info: Dumping cl_malloc item @ 0x805e350, 
> bucket address: 0x805e340
> Aug 28 11:42:38 hope lrmd: [17831]: info: Magic number: 0x636f2f62 
> reqsize=102, bucket=0, bucksize=32
> Aug 28 11:42:38 hope lrmd: [17831]: info: 68 61 6e 64 "hand"
> Aug 28 11:42:38 hope lrmd: [17831]: info: 6c 65 72 2d "ler-"
> Aug 28 11:42:38 hope lrmd: [17831]: info: 69 64 00 00 "id
> Aug 28 11:42:38 hope lrmd: [17831]: info: 39 00 00 00 "9
> Aug 28 11:42:38 hope lrmd: [17831]: info: ef be ed fe "���\376"
> Aug 28 11:42:38 hope lrmd: [17831]: info: 0c 00 00 00 "^L
> Aug 28 11:42:38 hope lrmd: [17831]: info: 00 00 00 00 "
> Aug 28 11:42:38 hope lrmd: [17831]: info: a0 16 06 08 "�^V^F^H"
> Aug 28 11:42:38 hope lrmd: [17831]: info: 00 00 00 00 "
> Aug 28 11:42:38 hope lrmd: [17831]: info: 40 00 00 00 "@
> Aug 28 11:42:38 hope lrmd: [17831]: info: 00 00 00 00 "
> Aug 28 11:42:38 hope lrmd: [17831]: info: 5a a5 5a a5 "Z�Z�"
> Aug 28 11:42:38 hope lrmd: [17831]: info: 00 00 00 00 "
> Aug 28 11:42:38 hope last message repeated 2 times
> Aug 28 11:42:38 hope lrmd: [17831]: info: 5a a5 5a a5 "Z�Z�"
> Aug 28 11:42:38 hope lrmd: [17831]: info: 5a a5 5a a5 "Z�Z�"
> Aug 28 11:42:38 hope lrmd: [17831]: info: 39 00 00 00 "9
> Aug 28 11:42:38 hope lrmd: [17831]: info: ef be ed fe "���\376"
> Aug 28 11:42:38 hope lrmd: [17831]: info: 20 00 00 00 "
> Aug 28 11:42:38 hope lrmd: [17831]: info: 00 00 00 00 "
> Aug 28 11:42:38 hope lrmd: [17831]: info: 00 00 00 00 "
> Aug 28 11:42:38 hope lrmd: [17831]: info: 30 84 05 08 "0�^E^H"
> Aug 28 11:42:38 hope lrmd: [17831]: info: f0 cd eb b7 "����"
> Aug 28 11:42:38 hope lrmd: [17831]: info: a8 23 08 08 "�#^H^H"
> Aug 28 11:42:38 hope lrmd: [17831]: info: 20 82 ea b7 " ���"
> Aug 28 11:42:38 hope lrmd: [17831]: info: f0 de 0f 08 "��^O^H"
> Aug 28 11:42:38 hope heartbeat: [17814]: WARN: 
> Exiting /usr/lib/heartbeat/lrmd -r process 17831 killed by signal 6 
> [SIGABRT - Abort].
> Aug 28 11:42:38 hope heartbeat: [17814]: ERROR: 
> Exiting /usr/lib/heartbeat/lrmd -r process 17831 dumped core

Could you please post the backtrace or, even better, open the
bugzilla for this incl. the logs and the configuration.

> Aug 28 11:42:38 hope heartbeat: [17814]: ERROR: Respawning 
> client "/usr/lib/heartbeat/lrmd -r":
> Aug 28 11:42:38 hope heartbeat: [17814]: info: Starting child 
> client "/usr/lib/heartbeat/lrmd -r" (0,0)
> Aug 28 11:42:38 hope heartbeat: [5438]: info: 
> Starting "/usr/lib/heartbeat/lrmd -r" as uid 0  gid 0 (pid 5438)
> Aug 28 11:42:38 hope lrmd: [5438]: info: G_main_add_SignalHandler: Added 
> signal handler for signal 15
> Aug 28 11:42:38 hope lrmd: [5438]: info: G_main_add_SignalHandler: Added 
> signal handler for signal 17
> Aug 28 11:42:38 hope lrmd: [5438]: WARN: Core dumps could be lost if multiple 
> dumps occur.
> Aug 28 11:42:38 hope lrmd: [5438]: WARN: Consider setting non-default value 
> in /proc/sys/kernel/core_pattern (or equivalent) for maximum supportability
> Aug 28 11:42:38 hope lrmd: [5438]: WARN: Consider 
> setting /proc/sys/kernel/core_uses_pid (or equivalent) to 1 for maximum 
> supportability
> Aug 28 11:42:38 hope lrmd: [5438]: info: G_main_add_SignalHandler: Added 
> signal handler for signal 10
> Aug 28 11:42:38 hope lrmd: [5438]: info: G_main_add_SignalHandler: Added 
> signal handler for signal 12
> Aug 28 11:42:38 hope lrmd: [5438]: info: Started.
> 
> 
> I then have to go though the same kill -9 to restart heartbeat.
> 
> 
> Thanks in advance.
> -- 
> Bret Baptist
> Senior Network Administrator
> [EMAIL PROTECTED]
> Internet Exposure, Inc.
> http://www.iexposure.com
> (612)676-1946 x17
> 
> Providing Internet Services since 1995
> Web Development ~ Search Engine Marketing ~ Web Analytics
> Network Security ~ On Demand Tech Support ~ E-Mail Marketing
> ------------------------------------------
> _______________________________________________
> Linux-HA mailing list
> [email protected]
> http://lists.linux-ha.org/mailman/listinfo/linux-ha
> See also: http://linux-ha.org/ReportingProblems
_______________________________________________
Linux-HA mailing list
[email protected]
http://lists.linux-ha.org/mailman/listinfo/linux-ha
See also: http://linux-ha.org/ReportingProblems

Reply via email to