Publication:
Labeling Turkish news stories with CRF

Loading...
Thumbnail Image

Institution Authors

Advisor

Department

Journal Title

Journal ISSN

Volume Title

Publisher

IEEE

Research Projects

Organizational Units

Journal Issue

Abstract

Drastically document increase in Web requires semantic web applications in order to lead the Web to its full potential. Extracting important phrases in a document facilitates finding expected information. In this paper, a new approach that is labeling the main subject, main predicate, main location and main date of an electronic document is introduced. The main subject label tells whom or what the document about. The main predicate label tells what the subject is or does. The main location label tells where the activities passed and the main date label tells when the document passed. With the help of this new methodology, extraction of not only high level description of the content, but also the attribute of a phrase in a document is provided. As experimental set, Turkish news stories are selected. To use as a training and test set, manual labeling is made by human annotators. Then, different models for each label are implemented to extract the labels automatically and they are compared to manually labeled results to evaluation process of this study.

Description

Journal or Series

2013 7th International Conference on Application of Information and Communication Technologies

ISSN

ISBN

Rights

Keywords

Citation

Collections

Endorsement

Review

Supplemented By

Referenced By

Related Patent

Related Goal

1
Görüntülenme
0
İndirme
Altmetric
Dimensions
PlumX Metrikleri
BIP! Indicators
Google Scholar
Scholar'da Ara ↗