PUMA
Istituto di Scienza e Tecnologie dell'Informazione     
Nottelmann H., Straccia U. Information retrieval and machine learning for probabilistic schema matching. In: Information Processing & Management, vol. 43 (3) pp. 552 - 576. Elsevier, 2007.
 
 
Abstract
(English)
Schema matching is the problem of finding correspondences (mapping rules, e.g. logical formulae) between heterogeneous schemas e.g. in the data exchange domain, or for distributed IR in federated digital libraries. This paper introduces a probabilistic framework, called sPLMap, for automatically learning schema mapping rules, based on given instances of both schemas. Different techniques, mostly from the IR and machine learning fields, are combined for finding suitable mapping candidates. Our approach gives a probabilistic interpretation of the prediction weights of the candidates, selects the rule set with highest matching probability, and outputs probabilistic rules which are capable to deal with the intrinsic uncertainty of the mapping process. Our approach with different variants has been evaluated on several test sets.
URL: http://scienceserver.cilea.it/pdflinks/07011415355512945.pdf
Subject Schema matching
H.3 Information Storage and Retrieval


Icona documento 1) Download Document PDF


Icona documento Open access Icona documento Restricted Icona documento Private

 


Per ulteriori informazioni, contattare: Librarian http://puma.isti.cnr.it

Valid HTML 4.0 Transitional