Abstract

This paper presents a method to calculate the semantic similarity with TongyiciCiLin and Word2vec. In the part of CiLin, the semantic similarity of words is calculated by using the distance of words as the main factor, the number of branches and the distance between branches as the fine-tuning parameters. In the part of Word2vec, this paper constructs a special Corpus based on movie review, and uses Word2vec model to calculate the semantic similarity of Chinese words. Then, the final semantic similarity is calculated by using the dynamic weighting strategy to fuse CiLin and Word2vec. The method makes full use of the semantic information of words in the knowledge base and Corpus. The experimental results show that the algorithm has better accuracy and more robust to domain sensitivity.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.