Knowledge and Information Systems

, Volume 2, Issue 1, pp 73–96

Database Integration Using Neural Networks: Implementation and Experiences

  • Wen-Syan Li
  • Chris Clifton
  • Shu-Yao Liu
Original Paper

DOI: 10.1007/s101150050004

Cite this article as:
Li, WS., Clifton, C. & Liu, SY. Knowledge and Information Systems (2000) 2: 73. doi:10.1007/s101150050004
  • 212 Downloads

Abstract.

Applications in a wide variety of industries require access to multiple heterogeneous distributed databases. One step in heterogeneous database integration is semantic integration: identifying corresponding attributes in different databases that represent the same real world concept. The rules of semantic integration can not be ‘pre-programmed’ since the information to be accessed is heterogeneous and attribute correspondences could be fuzzy. Manually comparing all possible pairs of attributes is an unreasonably large task. We have applied artificial neural networks (ANNs) to this problem. Metadata describing attributes is automatically extracted from a database to represent their ‘signatures’. The metadata is used to train neural networks to find similar patterns of metadata describing corresponding attributes from other databases. In our system, the rules to determine corresponding attributes are discovered through machine learning. This paper describes how we applied neural network techniques in a database integration problem and how we represent an attribute with its metadata as discriminators. This paper focuses on our experiments on effectiveness of neural networks and each discriminator. We also discuss difficulties of using neural networks for this problem and our wish list for the Machine Learning community.

Keywords: Artificial neural networks; Attribute correspondence identification; Database integration; Heterogeneous database

Copyright information

© Springer-Verlag London Limited 2000

Authors and Affiliations

  • Wen-Syan Li
    • 1
  • Chris Clifton
    • 2
  • Shu-Yao Liu
    • 3
  1. 1.C&C Research Laboratories, NEC USA, San Jose, USAUS
  2. 2.The MITRE Corporation, Bedford, USAUS
  3. 3.Oracle Corporation, Redwood Shores, USAUS