Measuring the similarity between implicit semantic relations from the Web

Danushka Bollegala*, Yutaka Matsuo, Mitsuru Ishizuka

*この研究の対応する著者

研究成果: Conference contribution

52 被引用数 (Scopus)

抄録

Measuring the similarity between semantic relations that hold among entities is an important and necessary step in various Web related tasks such as relation extraction, information retrieval and analogy detection. For example, consider the case in which a person knows a pair of entities (e.g. Google, You Tube), between which a particular relation holds (e.g. acquisition). The person is interested in retrieving other such pairs with similar relations (e.g. Microsoft, Powerset). Existing keyword-based search engines cannot be applied directly in this case because, in keyword-based search, the goal is to retrieve documents that are relevant to the words used in a query - not necessarily to the relations implied by a pair of words. We propose a relational similarity measure, using a Web search engine, to compute the similarity between semantic relations implied by two pairs of words. Our method has three components: representing the various semantic relations that exist between a pair of words using automatically extracted lexical patterns, clustering the extracted lexical patterns to identify the different patterns that express a particular semantic relation, and measuring the similarity between semantic relations using a metric learning approach. We evaluate the proposed method in two tasks: classifying semantic relations between named entities, and solving word-analogy questions. The proposed method outperforms all baselines in a relation classification task with a statistically significant average precision score of 0:74. Moreover, it reduces the time taken by Latent Relational Analysis to process 374 word-analogy questions from 9 days to less than 6 hours, with an SAT score of 51%. Copyright is held by the International World Wide Web Conference Committee (IW3C2).

本文言語English
ホスト出版物のタイトルWWW'09 - Proceedings of the 18th International World Wide Web Conference
ページ651-660
ページ数10
DOI
出版ステータスPublished - 2009
外部発表はい
イベント18th International World Wide Web Conference, WWW 2009 - Madrid
継続期間: 2009 4月 202009 4月 24

Other

Other18th International World Wide Web Conference, WWW 2009
CityMadrid
Period09/4/2009/4/24

ASJC Scopus subject areas

  • コンピュータ ネットワークおよび通信

フィンガープリント

「Measuring the similarity between implicit semantic relations from the Web」の研究トピックを掘り下げます。これらがまとまってユニークなフィンガープリントを構成します。

引用スタイル