'srmd executable error'? Has it been resolved?
Gerard Hynes <[email protected]> 20 May 2002 13:20:23 -0500
| Newsgroups | gmane.linux.failsafe |
|---|---|
| Message-ID | <[email protected]> |
Just got back to working with FailSafe after a couple of weeks on
another project.....
Anyway - grabbed the latest CVS from ftp.suse.com - nice work there
folks!! Compiles nicely on RH-7.2/GLIBC-2.2.4 and RH-7.3/GLIBC-2.2.5.
However - I am not faced with a small dilemma:
I don't use the GUI, as I'd like to script as much of this as possible.
I can do an RPM install, initialize the database, run a canned build
script which defines 2 machines - each with a public and private
interface, a shared IP address and a shared Apache resource. So far
so good. Currently (during this test phase) I am using SSH/STONITH
and that seems to work (albeit in a _really_ brutal fashion).
However - upon doing a haActivate - I get errors in the log file about
the IP_address script failing and a ''srmd executable error'' when doing
a haStatus.
I noted a similar thread recently in the mailing list. Has there been
a resolution to this?
Here's a snippet from haStatus -a ......
----- Cut Here -----
Cluster HA-CLUSTER:
Cluster state is ACTIVE.
Cluster Notify Cmd: "/bin/mail"
Cluster Notify Address: "fsafe_admin@localhost"
Node machine-b:
State of machine is UP.
Logical Machine Name: machine-b
Hostname: machine-b
Is FailSafe: true
Nodeid: 2
Reset type: powerCycle
System Controller: stonith
System Controller status: enabled
System Controller owner: machine-a
System Controller owner device: ssh
System Controller owner type: tty
ControlNet Ipaddr: 192.168.0.2
ControlNet HB: true
ControlNet Control: true
ControlNet Priority: 1
ControlNet Ipaddr: 10.8.0.66
ControlNet HB: true
ControlNet Control: true
ControlNet Priority: 2
Node machine-a:
State of machine is UP.
Logical Machine Name: machine-a
Hostname: machine-a
Is FailSafe: true
Nodeid: 1
Reset type: powerCycle
System Controller: stonith
System Controller status: enabled
System Controller owner: machine-b
System Controller owner device: ssh
System Controller owner type: tty
ControlNet Ipaddr: 192.168.0.1
ControlNet HB: true
ControlNet Control: true
ControlNet Priority: 1
ControlNet Ipaddr: 10.8.0.65
ControlNet HB: true
Resource_group RG1:
State: Online
Error: srmd executable error
Owner: machine-b
Failover Policy: IP-FAIL
Version: 1
Script: ordered
Attributes: Inplace_Recovery InPlace_Recovery
Controlled_Failback
Initial AFD: machine-a machine-b
Resources:
WebServer (type: Apache)
10.8.0.165 (type: IP_address)
Resource WebServer (type Apache):
State: Offline
Error: None
Owner: none
Flags: Resource is not locally monitored
port-number: 80
monitor-level: 1
default-page-location: /var/www/html/index.html
web-ipaddr: 10.8.0.165
server-root: /etc/httpd/conf
Resource dependencies
IP_address 10.8.0.165
Resource 10.8.0.165 (type IP_address):
State: Offline
Error: None
Owner: none
Flags: Resource is not locally monitored
BroadcastAddress: 10.8.0.255
interfaces: eth0:0
NetworkMask: 0xffffff00
No resource dependencies
Failover_policy IP-FAIL:
Version: 1
Script: ordered
Attributes: Inplace_Recovery InPlace_Recovery
Controlled_Failback
Initial AFD: machine-a machine-b
----- Cut here -----
Ideas anyone? I'll pull out the log files momentarily and dig
deeper - but the IP_address script seems to be failing badly.
Thanks in advance,
--
=[gh]=
[email protected] < remove the no-spam-o >
GPG Fingerprint = E944 5617 4FE6 C950 407C 0A72 187F E01A 1FC5 0A08
signature.asc
(application/pgp-signature, 232 B)
-----BEGIN PGP SIGNATURE----- Version: GnuPG v1.0.6 (GNU/Linux) Comment: For info see http://www.gnupg.org iD8DBQA86T5nGH/gGh/FCggRAjyOAJ4rR4vqCJ8qytkBjqEYSaBdlTeA4gCfZxMT mAKh9mDNqrYzAzmSQZ9W5n0= =URmc -----END PGP SIGNATURE-----