Deep Web Data Source Classification Based on Query Interface Context

Zilu Cui; Yuchen Fu

Found an issue? Give us feedback

https://doi.org/10.1...arrow_drop_down

https://doi.org/10.1109/iccis....

Article . 2012 . Peer-reviewed

Data sources: Crossref

https://dx.doi.org/10.1109/icc...

Article

Data sources: Microsoft Academic Graph

Deep Web Data Source Classification Based on Query Interface Context

descriptionPublicationkeyboard_double_arrow_right Article 01 Aug 2012Publisher:IEEEJournal:2012 Fourth International Conference on Computational and Information Sciences

Authors: Zilu Cui; Yuchen Fu;

doi: 10.1109/iccis.2012.117

Deep Web Data Source Classification Based on Query Interface Context

- Summary
- Metrics

Abstract

As the volume of information in the Deep Web grows, a Deep Web data source classification algorithm based on query interface context is presented. Two methods are combined to get the search interface similarity. One is based on the vector space. The classical TF-IDF statistics are used to gain the similarity between search interfaces. The other is to compute the two pages semantic similarity by the use of HowNet. Based on the K-NN algorithm, a WDB classification algorithm is presented. Experimental results show this algorithm generates high-quality clusters, measured both in terms of entropy and F-measure. It indicates the practical value of application.

Related Organizations

Soochow University
China (People's Republic of)

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	0
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

0

Average

Fields of Science

engineering and technology

electrical engineering, electronic engineering, information engineering

Fields of Science

engineering and technology

electrical engineering, electronic engineering, information engineering

Upload OA version

Are you the author of this publication? Upload your Open Access version to Zenodo!

It’s fast and easy, just two clicks!

uploadUpload now