Re: NIF segfault when using dirty schedulers

Steve Vinoski <[email protected]> Thu, 21 Jan 2016 07:24:58 -0500
Newsgroups gmane.comp.lang.erlang.bugs
Message-ID <CAO+zUOWFGHpW7OM97MoCDcgpunsmpKBdMw-fZZB7oGj-_o4bkg@mail.gmail.com>
--===============0051545446050968044==
Content-Type: multipart/alternative; boundary=001a11411c34e93d200529d7300b

--001a11411c34e93d200529d7300b
Content-Type: text/plain; charset=UTF-8

On Thu, Jan 21, 2016 at 12:26 AM, Paul Davis <[email protected]>
wrote:

> Hey all,
>
> I've recently run into a segfault while working with dirty schedulers.
> I managed to make a fairly concise reproducing test case at [1]. I
> included a stack trace at [2] from when the segfault occurs. This is
> definitely a racey segfault as well. I sometimes have to run `rebar
> eunit` a handful of times to trigger it.
>
> I'm not hugely familiar with all of the VM internals so I'm at a bit
> of a loss on where to start looking further. I did try and get rid of
> the requirement for eunit but I couldn't reproduce without it.
>
> This reproduces on both 17.5.6.4 where I found it and 18.2.2. I
> haven't tried master or anything of that nature.
>
> Let me know if there's anything else I can do to help debug this.
>
> Thanks,
> Paul
>
> [1] https://gist.github.com/davisp/1e71ec7f2f7a70d1b79c
> [2]
> https://gist.github.com/davisp/1e71ec7f2f7a70d1b79c#file-gdb_backtrace-txt


I'll have a look at your example, thanks for putting it together. Fro your
backtrace I'm guessing the issue is already fixed by available patches
which are not yet consolidated in any one branch or release.

--steve

--001a11411c34e93d200529d7300b
Content-Type: text/html; charset=UTF-8
Content-Transfer-Encoding: quoted-printable

<div dir=3D"ltr"><br><div class=3D"gmail_extra"><br><div class=3D"gmail_quo=
te">On Thu, Jan 21, 2016 at 12:26 AM, Paul Davis <span dir=3D"ltr">&lt;<a h=
ref=3D"mailto:[email protected]" target=3D"_blank">paul.joseph.da=
[email protected]</a>&gt;</span> wrote:<br><blockquote class=3D"gmail_quote" st=
yle=3D"margin:0 0 0 .8ex;border-left:1px #ccc solid;padding-left:1ex">Hey a=
ll,<br>
<br>
I&#39;ve recently run into a segfault while working with dirty schedulers.<=
br>
I managed to make a fairly concise reproducing test case at [1]. I<br>
included a stack trace at [2] from when the segfault occurs. This is<br>
definitely a racey segfault as well. I sometimes have to run `rebar<br>
eunit` a handful of times to trigger it.<br>
<br>
I&#39;m not hugely familiar with all of the VM internals so I&#39;m at a bi=
t<br>
of a loss on where to start looking further. I did try and get rid of<br>
the requirement for eunit but I couldn&#39;t reproduce without it.<br>
<br>
This reproduces on both 17.5.6.4 where I found it and 18.2.2. I<br>
haven&#39;t tried master or anything of that nature.<br>
<br>
Let me know if there&#39;s anything else I can do to help debug this.<br>
<br>
Thanks,<br>
Paul<br>
<br>
[1] <a href=3D"https://gist.github.com/davisp/1e71ec7f2f7a70d1b79c" rel=3D"=
noreferrer" target=3D"_blank">https://gist.github.com/davisp/1e71ec7f2f7a70=
d1b79c</a><br>
[2] <a href=3D"https://gist.github.com/davisp/1e71ec7f2f7a70d1b79c#file-gdb=
_backtrace-txt" rel=3D"noreferrer" target=3D"_blank">https://gist.github.co=
m/davisp/1e71ec7f2f7a70d1b79c#file-gdb_backtrace-txt</a></blockquote><div><=
br></div><div>I&#39;ll have a look at your example, thanks for putting it t=
ogether. Fro your backtrace I&#39;m guessing the issue is already fixed by =
available patches which are not yet consolidated in any one branch or relea=
se.</div><div><br></div><div>--steve</div></div></div></div>

--001a11411c34e93d200529d7300b--

--===============0051545446050968044==
Content-Type: text/plain; charset="us-ascii"
MIME-Version: 1.0
Content-Transfer-Encoding: 7bit
Content-Disposition: inline

_______________________________________________
erlang-bugs mailing list
[email protected]
http://erlang.org/mailman/listinfo/erlang-bugs

--===============0051545446050968044==--