And another thing with mailtrainer and thick threshold training

"Ger Hobbelt" <[email protected]>
Newsgroups gmane.mail.spam.crm114
Message-ID <[email protected]>
Howdy again, long time no see and all that.

See mailtrainer lines 860 and 915 (test_traingood/test_train_spam):

		eval /:@: :*:pr: < :*:thick_threshold: : /

                eval /:@: :*:pr: > (0 - :*:thick_threshold:) : /

Given the fact that today's classify thresholds are assymmetric and -
for 'good' at least - larger than the 'thick threshold' configured for
training when using mailtrainer, shouldn't these lines be changed to
use the basic classify (assymmetric) check thresholds? Or am I
reasoning in the wrong direction here and should I lift the thick
threshold to the MAX of spam_threshold and good_threshold as I want to
ensure that emails which have just been LEARNed, *do* pass the
post-check classify cycle at all times: with the current TRAIN thick
threshold at 5 and the CLASSIFY thresholds at -5/+10, many 'good'
trained items will be STUCK in the 'DON'T KNOW' section for ever.

And THAT I do not like. Or should I bow my head and live with this?


-- 
Met vriendelijke groeten / Best regards,

Ger Hobbelt

--------------------------------------------------
web: http://www.hobbelt.com/
 http://www.hebbut.net/
mail: [email protected]
mobile: +31-6-11 120 978
--------------------------------------------------

-------------------------------------------------------------------------
This SF.net email is sponsored by: Microsoft
Defy all challenges. Microsoft(R) Visual Studio 2008.
http://clk.atdmt.com/MRT/go/vse0120000070mrt/direct/01/
lmpx.com only provides a reader for public news (NNTP) servers. It is not affiliated with the servers or forums shown here and is not responsible for the content of articles, which is written by their respective authors.