Publication:
A Hybrid Method for Extracting Turkish Part-Whole Relation Pairs from Corpus

Loading...
Thumbnail Image

Date

Institution Authors

Advisor

item.page.editor

Editor

Department

Journal Title

Journal ISSN

Volume Title

Publisher

IEEE

DOI

Research Projects

Organizational Units

Journal Issue

Abstract

Extraction of various semantic relation pairs from different sources (dictionary definitions, corpus etc.) with high accuracy is one of the most popular topics in natural language processing (NLP). In this study, a hybrid method is proposed to extract Turkish part-whole pairs from corpus. Corpus statistics, WordNet similarities and Word2Vec word vector similarities are used together in this study. Firstly, initial part-whole seeds are prepared and by using these seeds part-whole patterns are extracted from corpus. For each pattern, a reliability score is calculated and reliable patterns are selected to produce new pairs from corpus. Various reliability scores are used for new pairs. To measure success of method, 19 target whole words are selected and average 83% (first 10 pairs), 74% (first 20 pairs), 68% (first 30 pairs) precisions are obtained, respectively.

Description

Journal or Series

2016 24TH SIGNAL PROCESSING AND COMMUNICATION APPLICATION CONFERENCE (SIU)

ISSN

ISBN

978-1-5090-1679-2

Rights

Citation

Collections

Endorsement

Review

Supplemented By

Referenced By

Related Patent

Related Goal

0

Views

0

Downloads