This work is licensed under a Creative Commons Attribution-NonCommercial 4.0 International License
|
||||||||
|
Paper Details
Paper Title
Short Text Similarity Understanding with Word Embeddings
Authors
  Rutuja Subhash Gadekar,  Prof. Bhagwan Kurhe
Abstract
Understanding short texts is crucial to many applications, but challenges abound. First, short texts do not always observe the syntax of a written language. As a result, traditional natural language processing tools, ranging from part-of-speech tagging to dependency parsing, cannot be easily applied. Second, short texts usually do not contain sufficient statistical signals to support many state-of-the-art approaches for text mining such as topic modeling. Third, short texts are more ambiguous and noisy, and are generated in an enormous volume, which further increases the difficulty to handle them. We argue that semantic knowledge is required in order to better understand short texts. In this work, we build a prototype system for short text understanding which exploits semantic knowledge provided by a well-known knowledgebase and automatically harvested from a web corpus. Our knowledge-intensive approaches disrupt traditional methods for tasks such as text segmentation, part-of-speech tagging, and concept labeling, in the sense that we focus on semantics in all these tasks. We conduct a comprehensive performance evaluation on real-life data. The results show that semantic knowledge is indispensable for short text understanding, and our knowledge-intensive approaches are both effective and efficient in discovering semantics of short texts.
Keywords- Topic Model, Short Texts, Word Embeddings
Publication Details
Unique Identification Number - IJEDR1803020Page Number(s) - 107-109Pubished in - Volume 6 | Issue 3 | July 2018DOI (Digital Object Identifier) -    Publisher - IJEDR (ISSN - 2321-9939)
Cite this Article
  Rutuja Subhash Gadekar,  Prof. Bhagwan Kurhe,   "Short Text Similarity Understanding with Word Embeddings", International Journal of Engineering Development and Research (IJEDR), ISSN:2321-9939, Volume.6, Issue 3, pp.107-109, July 2018, Available at :http://www.ijedr.org/papers/IJEDR1803020.pdf
Article Preview
|
|
||||||
|