Re: [suse-oracle] ocfs2 problem after sles 10 SP2 upgrade

gabi draghici <[email protected]>
Newsgroups gmane.linux.suse.oracle.general
Message-ID <[email protected]>
After an update to last kernel patch (2.6.16.60-0.37) it seems that problem
2 and 3 are solved !
Problem 1 is an oracle bug 8363126, as it's stated on metalink ! I guess we
have to wait for that one ...


Thanks,
Gabriel

On Mon, Apr 13, 2009 at 9:51 PM, Arun Singh <[email protected]> wrote:

> Yes, That's what I wrote #1 & #3 are kernel issue. I remember we fixed this
> in SP2 kernel update.
>
> Possibly #2 is also fixed in latest kernel update.
>
> Kernel update will solve this issue.
>
> -Arun
>
> >>> On 4/13/2009 at 11:32 AM, "Alexei_Roudnev" <
> [email protected]>
> wrote:
>  > Problem 1 is not OCFS relarted, it is ASYNC-IO related (the bug in
> SLES10
> > SP2 kernel). Just upgrade to the latest kernel.
> >
> >
> > ----- Original Message -----
> > From: "Arun Singh" <[email protected]>
> > To: <[email protected]>
> > Sent: Monday, April 13, 2009 9:31 AM
> > Subject: Re: [suse-oracle] ocfs2 problem after sles 10 SP2 upgrade
> >
> >
> >> Hello Draghici,
> >>
> >> #1 & #3 are caused by same issue, and most likely addressed in latest
> SP2
> >> kernel update (2.6.16.60-0.37).
> >>
> >> #2 - It appears to be OCFS2 related.
> >>
> >> Please apply latest kernel patch and If it doesn't help open a bug with
> >> Novell.
> >>
> >> Thanks,
> >> Arun
> >>
> >>>>> On 4/13/2009 at 6:53 AM, gabi draghici <[email protected]>
> wrote:
> >>> Hello to everyone !
> >>>
> >>> We had an sles 10 Sp1 , oracle rac 10.2.0.4,  ocfs  combination wich
> >>> worked
> >>> fine for about an year !
> >>> After we've done some testing we decided to upgrade to sles 10  Sp2 !
> >>> After upgrade we've notice 3 problems so far :
> >>>
> >>>    1. rman ( oracle tool for backup) doesn't work anymore from the
> local
> >>> fs
> >>> :
> >>>
> >>>  Recovery Manager: Release 10.2.0.4.0 - Production on Mon Apr 13
> 16:43:19
> >>> 2009
> >>>  Copyright (c) 1982, 2007, Oracle.  All rights reserved.
> >>>  RMAN-00571:
> ===========================================================
> >>>  RMAN-00569: =============== ERROR MESSAGE STACK FOLLOWS
> ===============
> >>>  RMAN-00571:
> ===========================================================
> >>>  RMAN-00554: initialization of internal recovery manager package failed
> >>>  RMAN-03000: recovery manager compiler component initialization failed
> >>>  RMAN-06001: error parsing job step library
> >>>  RMAN-01006: error signalled during parse
> >>>  RMAN-00600: internal error, arguments [8083] [] [] [] []
> >>>  LFI-00005: Free some memory failed in lfibrdt().
> >>>  LFI-00004: Call to lfibgl() failed.
> >>>
> >>>    2. in /var/log/messages we have :
> >>>
> >>>
> >>> Apr 13 13:59:03 nod2 kernel: Call Trace:
> >>> <ffffffff8016357a>{remove_from_page_cache+49}
> >>> Apr 13 13:59:03 nod2 kernel:
> >>> <ffffffff801695b4>{truncate_complete_page+53}
> >>> <ffffffff8016965e>{truncate_inode_pages_range+159}
> >>> Apr 13 13:59:03 nod2 kernel:
> >>> <ffffffff884af364>{:ocfs2:ocfs2_data_convert_worker+202}
> >>> Apr 13 13:59:03 nod2 kernel:
> >>> <ffffffff884ad602>{:ocfs2:ocfs2_downconvert_thread+1190}
> >>> Apr 13 13:59:03 nod2 kernel:
> >>> <ffffffff80148092>{autoremove_wake_function+0}
> >>> <ffffffff884ad15c>{:ocfs2:ocfs2_downconvert_thread+0}
> >>> Apr 13 13:59:03 nod2 kernel:
> >>> <ffffffff80147c88>{keventd_create_kthread+0}
> >>> <ffffffff80147f50>{kthread+236}
> >>> Apr 13 13:59:04 nod2 kernel:        <ffffffff8010bed2>{child_rip+8}
> >>> <ffffffff80147c88>{keventd_create_kthread+0}
> >>> Apr 13 13:59:04 nod2 kernel:        <ffffffff80147e64>{kthread+0}
> >>> <ffffffff8010beca>{child_rip+0}
> >>> Apr 13 13:59:04 nod2 kernel: Badness in
> __remove_from_page_cache_nocheck
> >>> at
> >>> mm/filemap.c:122
> >>>
> >>> The only thing we found about that is a novell document, id = 7000562
> >>> wich
> >>> resolution is "please open a service request ... ".
> >>>
> >>>
> >>>  3. in oracle's alert.log we found that arch process can't write :
> >>>
> >>>
> >>>
> >>> ARC1: Encountered disk I/O error 19502
> >>> Mon Apr 13 13:02:36 2009
> >>> ARC1: Closing local archive destination LOG_ARCHIVE_DEST_1:
> >>> '/oracle/PRD/oraarch/PRDarch/1_12508_639435084.dbf' (error 19502)
> >>>  (PRD001)
> >>> Mon Apr 13 13:02:36 2009
> >>> Errors in file /oracle/PRD/saptrace/background/prd001_arc1_9381.trc:
> >>> ORA-19502: write error on file
> >>> "/oracle/PRD/oraarch/PRDarch/1_12508_639435084.dbf", blockno 32769
> >>> (blocksize=512)
> >>> ORA-27061: waiting for async I/Os failed
> >>> Linux-x86_64 Error: 5: Input/output error
> >>> Additional information: -1
> >>> Additional information: 1048576
> >>>
> >>>
> >>> This one we solved by disabling the disk_asynch_io and
> >>> filesystemio_options
> >>> ! The efect is that performance is affected (but at least the arch
> >>> process is ok now );
> >>>
> >>>
> >>> Any help is appreciated !
> >>>
> >>>
> >>> Draghici Gabriel
> >>> database administrator
> >>
> >> _______________________________________________
> >> suse-oracle mailing list
> >> [email protected]
> >> http://listx.novell.com/mailman/listinfo/suse-oracle
> >>
> _______________________________________________
> suse-oracle mailing list
> [email protected]
> http://listx.novell.com/mailman/listinfo/suse-oracle
>

_______________________________________________
suse-oracle mailing list
[email protected]
http://listx.novell.com/mailman/listinfo/suse-oracle
lmpx.com only provides a reader for public news (NNTP) servers. It is not affiliated with the servers or forums shown here and is not responsible for the content of articles, which is written by their respective authors.