Re: NOT SPAM: a request for information!

"bpallen777" <[email protected]>
Newsgroups gmane.comp.infodesign.facetedclassification
Message-ID <[email protected]>
Bob- The amount of work depends on how much metadata is available to 
the application developer, and how it lends itself to use in a 
faceted retrieval scheme. One observation to be made based on our 
experience is that there is far more metadata to be leveraged for 
this kind of application than people tend to assume.

In many commercial applications, we find that a significant amount 
of effort has already been spent in record classification by 
database designers and administrators. This existing metadata can be 
leveraged in a faceted retrieval system. For example, in an 
ecommerce catalog, products are often already classified by product 
type/category, target customer, price range, etc. In our system, we 
simple provide a way to automatically connect to relational tables 
containing such data, transform it into an RDF representation, and 
then automatically generate a navigation interface on top of the 
unified representation. In this way, a usable faceted search system 
can be generated in a few man-hours or less.

We also find simple Dublin Core metadata, which is readily available 
through many content management systems, can be fruitfully exploited 
to provide faceted search over tagged content. While this is not a 
faceted classification scheme in the traditional sense of a set of 
facets crafted for a specific set of end-user retrieval needs, it 
can nevertheless provide more usable navigation than is possible 
with traditional free text or parametric retrieval approaches. 
Coupled with named entity and noun phrase extraction techniques, 
this can lead to a point of departure for crafting a more 
sophisticated faceted classification scheme, while in the short term 
providing real benefit to users.

With the proliferation of XML-based approaches to content management 
(particularly those associated with weblog management systems), the 
amount of metadata that can be leveraged to build faceted schemes 
will continue to expand. Again, in our system, one transforms the 
metadata available in XML feeds and/or repositories into an RDF 
representation and continues from there.

In many cases, however, the intellectual effort of creating a 
classification scheme, specifying facets and providing a framework 
for a combination of human- and machine-generated metadata creation 
and harvesting will need to be performed. However, our intent is to 
provide tools that will, through the use of standards such as RDF, 
provide a framework so that such efforts can be easily shared, 
reused and propagated through the agency of the Semantic Web.  Our 
goal is to allow information architects to write RDF to create 
working navigation systems in precisely the same way web page 
designers write HTML to create web pages. In doing so, we can 
make "view source" work for information architectures in a manner 
analogous to the way it works for web pages today. 

For example, if you go to the DC-2002 example cited in my previous 
post, you'll see an "RDF Metadata" button in the lower left hand 
corner of the page. Clicking on that link will show you the source 
RDF that defines the running navigation application. It contains a 
simple controlled vocabulary using Z39.19 standards, a temporal 
hierarchy, a simple classification scheme for conference events, and 
the presentation metadata records. This took about a day to put 
together based on the raw material on the conference website. By 
providing the RDF representation directly, it now becomes possible 
for others to reuse the scheme to apply it to other collections and 
extend it for other uses.

In summary, there are a number of tools and approaches that we use 
to assist in the classification process:

- Extract metadata directly from relational databases and XML feeds;
- Support incremental development of a faceted classification from 
an initial set of Dublin Core metadata
- Use named entity and noun phrase extraction tools to create 
subjects from unstructured content;
- Author, aggregate and reuse existing RDF representations of 
vocabularies and classification schemes

regards, BPA


------------------------ Yahoo! Groups Sponsor ---------------------~-->
Get A Free Psychic Reading! Your Online Answer To Life's Important Questions.
http://us.click.yahoo.com/Lj3uPC/Me7FAA/ySSFAA/0bmwlB/TM
---------------------------------------------------------------------~->

----------
Thanks for playing. To unsubscribe from this group, send an email to:
[email protected]

 

Your use of Yahoo! Groups is subject to http://docs.yahoo.com/info/terms/
lmpx.com only provides a reader for public news (NNTP) servers. It is not affiliated with the servers or forums shown here and is not responsible for the content of articles, which is written by their respective authors.