Semantics, Knowledge and Grid, International Conference on (2012)
Beijing, TBD, China China
Oct. 22, 2012 to Oct. 24, 2012
DOI Bookmark: http://doi.ieeecomputersociety.org/10.1109/SKG.2012.22
Faceted search on web pages needs exact facets. However, it is difficult to extract facets exactly from web pages because the web pages are unstructured and lack of facet information. Therefore, facet extraction is a key to faceted search. This paper proposed a method of extracting facets automatically from unstructured web pages to improve the faceted search on web. The Multidimensional Semantic Index (MDSI) of web pages is constructed by mining all kinds of semantic relations among the words from web pages, which creates a semantic-rich index for web pages. In MDSI, the differently dimensional semantic indexes are bridged by mining the semantic mapping between them. Based on the MDSI of web pages, the facets are extracted by analyzing semantic mapping relations in MDSI. To validate the effect of the proposed method, two datasets are constructed and the experimental results show that the proposed method is feasible and comparatively precise.
facet extraction, faceted search, multidimensional semantic index, semantic mapping
X. Wei, X. Luo and Q. Li, "Automatic Facet Extraction Based on Multidimensional Semantic Index," 2012 Eighth International Conference on Semantics, Knowledge and Grids (SKG 2012)(SKG), Beijing, 2012, pp. 64-71.