Re: QA failures - holding the release
"Ken McDonell" <[email protected]>
| Newsgroups | gmane.comp.sysutils.pcp |
|---|---|
| Message-ID | <[email protected]> |
I've been holding off my QA report because things have generally not been good since I moved the QA farm ... I am still not comfortable that I have eliminated all of the failures attributable to the configuration change, but there are (comparatively) lots of failures. The report for all but the "onesies and twosies" is appended at the end of this mail. I'll comment on Mark's ones, but I'm seeing a somewhat larger failure landscape. On 14/11/16 12:25, Mark Goodwin wrote: > I've not yet been able to resolve the following QA failures on F24/x86_64. > Some just need some QA work (e.g. the nfsv4 related), others need better > filtering. But some are not understood (by me anyway). I'm concentrating > on the pmrep segfaults before tackling any others for now. > > Failures: 042 216 232 366 651 732 755 766 782 798 895 1069 > > 042 no values available for proc.memory.rss in busybox container > (other metrics are OK) Passes everywhere for me. > 216 network metrics out of range (0..0). Issue with version of netstat > on F24? >> network.icmp.inerrors = 152 out of range 0..0 >> network.icmp.outechos = 612 out of range 0..0 >> network.tcp.outsegs = 6703507 out of range 0..0 >> network.tcp.retranssegs = 77259 out of range 0..0 >> network.tcp.inerrs = 7440 out of range 0..0 >> network.udp.noports = 226 out of range 0..0 Passes everywhere for me, including my F24 host. > 232 torture indom, fetch != pmGetInDom for nfs4.client.reqs > nfs4.client.reqs: > number of instances from unprofiled fetch (0) != that for pmGetInDom (56) A couple of failures for me, so in the "onesies and twosies" group. > 366 logconf for jbd2 > $ rpm -qf $(pwd)/jbd2 > file /var/lib/pcp/config/pmlogconf/filesystem/jbd2 is not owned by any > package > ... hmm so how did it get there? Left over from a previous install? No help on this one I'm afraid. > 651 sample PMDA IPC protocol failure A couple of failures for me, so in the "onesies and twosies" group. > 732 nfs4.server.reqs has 11 additional entries on fc24 (allocate, copy, > copy-notify, etc) 4 hosts showing the same symptom. > 755 no apache agent? Passes for me everywhere it is run. > 766 > pcp://10.64.152.4:44321 <http://10.64.152.4:44321> > > pcp://10.64.152.4:44321 <http://10.64.152.4:44321> Passes for me everywhere it is run. > 782 pmwebd not terminating in a timely manner .. forced to terminate) Passes for me everywhere it is run. > 798 nfs value filtering? Lukas has committed a fix for this that so far I've verified works on two QA hosts. > 895 oops, 100 archive records not between the expected range of 300 and 1500 All my failures here are a pduread() interrupt diagnostic that probably needs filtering out (I'll fix that). > 1069 python, pmrep segfaults > $ pmrep -s 1 --archive archives/20130706 -z -O 30m > disk.dev.read,,"'sda','sdb'",,,16 > d.d.read d.d.read > sda sdb > count/s count/s > N/A N/A > Segmentation fault (core dumped) > must be segfaulting somewhere in the C libraries? > I have one segfault, and all the rest seem to have TypeError: non-empty format string passed to object.__format__ in the failure traceback. ==== QA Summary ==== Date Run Pass Fail Nrun Host 2016-11-06 779 772 7 62|bo|bozo PCP 3.11.4 x86_64 Ubuntu 16.04 Daily runs, but no QA |bl|bozo-laptop PCP 3.11.4 i686 LinuxMint 15 Daily runs, but no QA |bv|bozo-vm PCP 3.11.5 x86_64 Debian 8.5 Daily runs, but no QA |fu|fuji PCP 3.11.4 i386 Darwin 10.8.0 Daily runs, but no QA |gr|grundy.sgi.com|grundy.sgi.com 2016-11-11 869 865 4 83|00|vm00 PCP 3.11.4 x86_64 Ubuntu 12.04 2016-11-11 773 766 7 67|01|vm01 PCP 3.11.4 i686 Ubuntu 15.10 2016-11-11 860 849 11 92|02|vm02 PCP 3.11.4 i686 openSUSE 13.2 2016-11-04 333 330 3 18|03|vm03 PCP 3.11.4 x86_64 Fedora 24 2016-11-09 766 758 8 186|04|vm04 PCP 3.11.4 i686 CentOS 5.11 2016-11-12 865 860 5 87|05|vm05 PCP 3.11.4 x86_64 Gentoo 2.2 2016-11-12 61 61 0 4|06|vm06 PCP 3.11.4 amd64 FreeBSD 10.2-RELEASE 2016-11-12 870 863 7 82|07|vm07 PCP 3.11.4 x86_64 Debian stretch/sid 2016-11-14 891 885 6 61|08|vm08 PCP 3.11.4 x86_64 CentOS Linux7.2.1511 2016-11-14 61 58 3 4|09|vm09 PCP 3.11.4 i386 NetBSD 6.1.5 2016-11-12 61 61 0 4|10|vm10 PCP 3.11.4 i386 FreeBSD 9.3-RELEASE-p30 2016-11-12 868 865 3 84|11|vm11 PCP 3.11.5 i686 Debian stretch/sid 2016-11-12 887 884 3 65|12|vm12 PCP 3.11.5 i686 Fedora 22 Daily runs, but no QA |13|vm13 PCP 3.9.1 i86pc OpenIndiana Development oi_151.1.4 2016-11-13 884 880 4 68|14|vm14 PCP 3.11.5 x86_64 CentOS6.7 2016-11-14 815 811 4 137|15|vm15 PCP 3.11.4 x86_64 Slackware "14.2" 2016-11-13 882 876 6 70|18|vm18 PCP 3.11.4 x86_64 LinuxMint 17.3 2016-11-10 62 51 11 4|19|vm19 PCP 3.11.4 x86_64 openSUSE 12.2 2016-11-13 884 876 8 68|20|vm20 PCP 3.11.4 x86_64 Ubuntu 14.04 2016-11-13 865 862 3 87|21|vm21 PCP 3.11.4 i686 Debian 7.10 2016-11-13 885 882 3 67|22|vm22 PCP 3.11.4 x86_64 Fedora 19 2016-11-14 888 885 3 64|23|vm23 PCP 3.11.4 i686 Fedora 20 2016-11-14 333 328 5 18|24|vm24 PCP 3.11.4 i686 openSUSE 13.1 2016-11-14 761 754 7 191|25|vm25 PCP 3.11.4 x86_64 CentOS 5.11 2016-11-12 887 884 3 65|26|vm26 PCP 3.11.4 x86_64 Fedora 21 2016-11-13 877 868 9 75|27|vm27 PCP 3.11.5 x86_64 Ubuntu 15.04 2016-11-14 881 875 6 71|28|vm28 PCP 3.11.5 x86_64 RHEL Server 6.8 2016-11-13 888 877 11 64|29|vm29 PCP 3.11.4 x86_64 RHEL Server 7.2 2016-11-14 886 881 5 66|30|vm30 PCP 3.11.4 x86_64 SUSE SLES12 SP0 2016-11-15 886 881 5 66|31|vm31 PCP 3.11.4 x86_64 Fedora 23 2016-11-14 74 74 0 5|32|vm32 PCP 3.11.5 amd64 FreeBSD 11.0-CURRENT 2016-11-14 64 62 2 1|33|vm33 PCP 3.11.4 amd64 OpenBSD 5.8 2016-11-14 877 872 5 75|34|vm34 PCP 3.11.4 x86_64 Arch Linux Summary: 22523 run, 167 failed (0.74%) ==== QA Failure (X) and Not Run or Skipped (-) Map ==== Host bo 00 01 02 03 04 05 07 08 09 11 12 14 15 18 19 20 21 22 23 24 25 26 27 28 29 30 31 33 34 Test %fail Test QA groups 884 61% X X X X - X X X X - X X X X - X X X X - X X X X - 884 libpcp_web 003 52% X X X - X X X X X X X X X X X X X - X 003 pdu pmcd mem_leak 895 24% - X - X X X - X X - X X - 895 pmlogger 1069 24% - - X - - X X - X X - X - - X - X 1069 pmrep python timezone 798 18% - X - X - - X X X - X 798 pmda.nfsclient 069 15% - X - X X X X - 069 pmcd pmval 083 15% - X - - X X X X - 083 pmlc pmlogger compat 732 15% - - X - - X X X - X 732 pmda.linux 778 15% X X - - - - - - - - - - - - X - X - - - - - - X - - - - - - 778 pmda.postgresql pmie 232 12% - X X X X 232 libpcp 365 12% X - X - - X X - 365 pmcd 888 12% X - - X - - X X - 888 pmda.linux 1072 12% - - - - X - X - X - - X - 1072 pmrep python archive 243 9% X X X - - - 243 pmcd pmprobe 651 9% - X - X - - X - 651 pmproxy 835 9% X - - - - - - - - - - - - - X - X - - - - - - - - - - - - - 835 pmda.memcache 964 9% X X - - - - - - X - 964 pmcd 1038 9% - - - - X - - X - - X - 1038 pmrep archive multi-archive Host bo 00 01 02 03 04 05 07 08 09 11 12 14 15 18 19 20 21 22 23 24 25 26 27 28 29 30 31 33 34 Test %fail Test QA groups 1071 9% - - - - X - - X - - X - 1071 pmrep python -=-=-=-=-=-=-=-=-=-=-=- pcp mailing list [email protected] https://groups.io/g/pcp/messages -=-=- Groups.io Links: You receive all messages sent to this group. View/Reply Online (#14697): https://groups.io/g/pcp/message/14697 View All Messages In Topic (4): https://groups.io/g/pcp/topic/3119244 Mute This Topic: https://groups.io/mt/3119244?uid=174580 New Topic: https://groups.io/g/pcp/post Change Your Subscription: https://groups.io/g/pcp/editsub?uid=174580 Group Home: https://groups.io/g/pcp Contact Group Owner: [email protected] Terms of Service: https://groups.io/static/tos Unsubscribe: https://groups.io/g/pcp/leave/354243/563757577/xyzzy -=-=-=-=-=-=-=-=-=-=-=-