loading...
 This Article 
   
 Share 
   
 Bibliographic References 
   
 Add to: 
 
Digg
Furl
Spurl
Blink
Simpy
Google
Del.icio.us
Y!MyWeb
 
 Search 
   
8th International Symposium on Parallel Architectures,Algorithms and Networks (ISPAN'05)
Las Vegas, Nevada, USA
December 07-December 09
ISBN: 0-7695-2509-1
Mauricio Hernandez, IBM Almaden Research Center
Lucian Popa, IBM Almaden Research Center
Howard Ho, IBM Almaden Research Center
Felix Naumann, Humboldt University at Berlin

Information integration typically requires the construction of complex artifacts like federated databases, ETL scripts, data warehouses, applications for accessing multiple data sources, and applications that ingest or publish XML. For many companies, it is one of the most complicated IT tasks they face today. To reduce the overall cost, intelligent tools are needed to simplify this difficult task.

Clio is a semi-automatic tool for schema mapping and data integration developed at IBM Almaden Research Center over the past few years. It takes source and target schemas as input, which may describe relational or XML data models. Via a graphical SchemaViewer, a user can then interactively specify attribute correspondences between the source and target schemas. An AttributeMatcher component helps suggest such correspondences, based on the similarity of both attribute names and attribute values. Once the user has specified correspondences, Clio generates SQL, SQL/XML, XQuery or XSLT on the fly to implement the specified transformation, which is guaranteed to produce output data that conforms to the target schema. In this talk, we will first describe and demonstrate some basic features of Clio. In particular, we will describe the abstracted problems and the algorithms behind the AttributeMatcher component. Then, we will describe additional research problems abstracted from the area of schema mapping and information integration, with an emphasis on graph algorithms and issues on scalability and parallelism.

Citation:
Mauricio Hernandez, Lucian Popa, Howard Ho, Felix Naumann, "Clio: A Schema Mapping Tool for Information Integration," ispan, pp.11, 8th International Symposium on Parallel Architectures,Algorithms and Networks (ISPAN'05), 2005
Usage of this product signifies your acceptance of the Terms of Use.