Re: size of variables in hyperspace?
[email protected] (Bill Yerazunis) Wed, 10 Jun 2009 09:18:23 -0400 (EDT)
| Newsgroups | gmane.mail.spam.crm114 |
|---|---|
| Message-ID | <20090610131823.9C87B3DE2B6@starbuck> |
From: Thomas Michael Hagen <[email protected]> > Ouch... in fact, I'd be worried that the boilerplate around it (especially > the javascript) may be including "telltales" that are dominating > the decision process. does that mean hyperspace does not limit itself to storing the unique features of each category, like osbf does? Correct. Hyperspace does not coalesce multiple copies of the same feature from different examples. The features are independent (because they represent points in a 4-billion-dimension hyperspace) usually, in machine learning, it is a good idea to include as much information as possible, including things that would only confuse a human being... Depends - if it's noise (i.e. uncorrellated with the desired result and a high standard deviation) or common ( uncorrellated, but with a low standard deviation) then getting rid of it is a good thing. But yes, if it is correlated with the desired result, you should keep it. The easy way to know for sure is to just do the experiment... - Bill Yerazunis ------------------------------------------------------------------------------ Crystal Reports - New Free Runtime and 30 Day Trial Check out the new simplified licensing option that enables unlimited royalty-free distribution of the report engine for externally facing server and web deployment. http://p.sf.net/sfu/businessobjects