J4 ›› 2010, Vol. 48 ›› Issue (05): 811-816.
Previous Articles Next Articles
ZHANG Dongna, ZHOU Chunguang, LIU Yanbin, GUO Dongwei
Received:
Online:
Published:
Contact:
Abstract:
We first proposed a new method calculating semantic similarity parameter information content. The new algorithm is based on the concept semantic information in the knowledge base called WordNet and the probability in the corpus called selfinformation. Then, considering the existing algorithms are all domainrelated and the calculating processes are complicated, we proposed a universal method based on corpus statistics and WordNet calculating semantic similarity which can be used in information extraction, information retrieval, document clustering and ontology learning. The proposed method makes a substantial improvement experimenting on the benchmark data setR&B concept pairs.
Key words: semantic similarity of concepts, Brown corpus, information content method
CLC Number:
ZHANG Dong-Na, ZHOU Chun-Guang, LIU Pan-Bin, GUO Dong-Wei. A Semantic Similarity Computing Approach Based onWordNet and Corpus Statistics[J].J4, 2010, 48(05): 811-816.
0 / / Recommend
Add to citation manager EndNote|Reference Manager|ProCite|BibTeX|RefWorks
URL: https://xuebao.jlu.edu.cn/lxb/EN/
https://xuebao.jlu.edu.cn/lxb/EN/Y2010/V48/I05/811
Cited