A machine learning approach for correcting the errors of a Treebank

Zarei, Farzaneh; Faili, Hesham; Mirian, Maryam

Volume 12, Issue 3 (12-2015) JSDP 2015, 12(3): 99-108 | Back to browse issues page

Mendeley

Zotero

RefWorks

Zarei F, Faili H, Mirian M. A machine learning approach for correcting the errors of a Treebank . JSDP 2015; 12 (3) :99-108
URL: http://jsdp.rcisp.ac.ir/article-1-221-en.html

A machine learning approach for correcting the errors of a Treebank

Farzaneh Zarei ^*

, Hesham Faili

, Maryam Mirian

Abstract: (5633 Views)

The Treebank is one of the most useful resources for supervised or semi-supervised learning in many NLP tasks such as speech recognition, spoken language systems, parsing and machine translation. Treebank can be developded in different ways that could be, generally, categorized in manually and statistical approaches. While the resulted Treebank in each of these methods has the annotation error, one which accomplished by statistical method has much more errors than the other. Error in Treenabanks causes that they are not useful anymore. In this paper an statistical method is proposed which aims to correct the errors in a specific English LTAG-Treebank. The proposed method was applied to a automatically generated Treebank and an improvement from 68% to 79% respect to F-measure is retrieved.

Full-Text [PDF 1767 kb] (1615 Downloads)

Type of Study: Research | Subject: Paper
Received: 2014/03/8 | Accepted: 2015/08/4 | Published: 2016/01/4 | ePublished: 2016/01/4

Send email to the article author

Rights and permissions
	This work is licensed under a Creative Commons Attribution-NonCommercial 4.0 International License.

Signal and Data Processing

Vote