Re: [Nutch-dev] plugin proposal

Stefan Groschupf Thu, 20 May 2004 03:29:47 -0700

* do you plan to store metadata in inverted lists as well, (which currently would translate into adding new arbitrary fields in Lucene's Document)? This would be very useful in my scenario - I'm doing language detection and key-phrase extraction to enhance the index, and I'd love to store this information in the index itself to avoid the need for separate storage.

+1 that i was asking with how to story dynamically meta data for each page, since i wish to do something similar like key phrase extraction. I would be interested to hear how you extract you key phrases?

Do you know Kea? It is interesting but does not scale at all. http://www.text-mining.org/index.jsp? action=showDocumentFromFolder&folderPK=789&documentPK=2500&template=docD etail

Stefan

---------------------------------------------------------------
open technology:   http://www.media-style.com
open source:           http://www.weta-group.net
open discussion:    http://www.text-mining.org

------------------------------------------------------- This SF.Net email is sponsored by: Oracle 10g Get certified on the hottest thing ever to hit the market... Oracle 10g. Take an Oracle 10g class now, and we'll give you the exam FREE. http://ads.osdn.com/?ad_id=3149&alloc_id=8166&op=click _______________________________________________ Nutch-developers mailing list [EMAIL PROTECTED] https://lists.sourceforge.net/lists/listinfo/nutch-developers

Re: [Nutch-dev] plugin proposal

Reply via email to