A methodology for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level following the PropBank-VerbNet model

In this article we describe the methodology developed for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level that follows the PropBank-VerbNet model. The methodology presented is the product of detailed theoretical study of the semantic nature of verbs in Basque...

Descripción completa

Detalles Bibliográficos
Autores: Estarrona Ibarloza, Ainara, Aldezabal Roteta, Izaskun, Díaz de Ilarraza Sánchez, Arantza, Aranzabe Urruzola, María Jesús
Tipo de recurso: artículo
Fecha de publicación:2015
País:España
Institución:Universidad del País Vasco
Repositorio:Addi. Archivo Digital para la Docencia y la Investigación
OAI Identifier:oai:addi.ehu.eus:10810/71090
Acceso en línea:http://hdl.handle.net/10810/71090
Access Level:acceso abierto
id ES_7f256c8d1d1c1744db6257920f2e2db3
oai_identifier_str oai:addi.ehu.eus:10810/71090
network_acronym_str ES
network_name_str España
repository_id_str
spelling A methodology for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level following the PropBank-VerbNet modelEstarrona Ibarloza, AinaraAldezabal Roteta, IzaskunDíaz de Ilarraza Sánchez, ArantzaAranzabe Urruzola, María JesúsIn this article we describe the methodology developed for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level that follows the PropBank-VerbNet model. The methodology presented is the product of detailed theoretical study of the semantic nature of verbs in Basque and of their similarities and differences with verbs in other languages. As part of the proposed methodology, we are creating a Basque lexicon on the PropBank-VerbNet model that we have named the Basque Verb Index (BVI). Our work thus dovetails with the general trend toward building lexicons from tagged corpora that is clear in work conducted for other languages. EPEC-RolSem and BVI are two important resources for the computational semantic processing of Basque; as far as the authors are aware, they are also the first resources of their kind developed for Basque. In addition, each entry in BVI is linked to the corresponding verb-entry in well-known resources like PropBank, VerbNet, WordNet, FrameNet, and Levin’s classification. We have also implemented several automatic processes to aid in creating and annotating the BVI, including processes designed to facilitate the task of manual annotation.This research has been supported by the University of the Basque Country (GIU09/19), the Basque Government (IXA group, Research Group of type A (2010–2015) (IT344-10), EUS-SRL project (S-PE11UN098), and Berbatek project (IE09-262)), and The Ministry of Science and Innovation of the Spanish Government (EPEC-RolSem project (FFI2008-02805-E/FILO) (Complementary Action) and AncoraNet project (FFI2009-06497-E)).Oxford University Press202520252015info:eu-repo/semantics/articleapplication/pdfhttp://hdl.handle.net/10810/71090reponame:Addi. Archivo Digital para la Docencia y la Investigacióninstname:Universidad del País VascoIngléshttps://doi.org/10.1093/llc/fqv001info:eu-repo/semantics/openAccess© The Author 2015. Published by Oxford University Press on behalf of EADH.oai:addi.ehu.eus:10810/710902026-06-18T09:23:17Z
dc.title.none.fl_str_mv A methodology for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level following the PropBank-VerbNet model
title A methodology for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level following the PropBank-VerbNet model
spellingShingle A methodology for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level following the PropBank-VerbNet model
Estarrona Ibarloza, Ainara
title_short A methodology for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level following the PropBank-VerbNet model
title_full A methodology for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level following the PropBank-VerbNet model
title_fullStr A methodology for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level following the PropBank-VerbNet model
title_full_unstemmed A methodology for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level following the PropBank-VerbNet model
title_sort A methodology for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level following the PropBank-VerbNet model
dc.creator.none.fl_str_mv Estarrona Ibarloza, Ainara
Aldezabal Roteta, Izaskun
Díaz de Ilarraza Sánchez, Arantza
Aranzabe Urruzola, María Jesús
author Estarrona Ibarloza, Ainara
author_facet Estarrona Ibarloza, Ainara
Aldezabal Roteta, Izaskun
Díaz de Ilarraza Sánchez, Arantza
Aranzabe Urruzola, María Jesús
author_role author
author2 Aldezabal Roteta, Izaskun
Díaz de Ilarraza Sánchez, Arantza
Aranzabe Urruzola, María Jesús
author2_role author
author
author
description In this article we describe the methodology developed for the semiautomatic annotation of EPEC-RolSem, a Basque corpus labeled at predicate level that follows the PropBank-VerbNet model. The methodology presented is the product of detailed theoretical study of the semantic nature of verbs in Basque and of their similarities and differences with verbs in other languages. As part of the proposed methodology, we are creating a Basque lexicon on the PropBank-VerbNet model that we have named the Basque Verb Index (BVI). Our work thus dovetails with the general trend toward building lexicons from tagged corpora that is clear in work conducted for other languages. EPEC-RolSem and BVI are two important resources for the computational semantic processing of Basque; as far as the authors are aware, they are also the first resources of their kind developed for Basque. In addition, each entry in BVI is linked to the corresponding verb-entry in well-known resources like PropBank, VerbNet, WordNet, FrameNet, and Levin’s classification. We have also implemented several automatic processes to aid in creating and annotating the BVI, including processes designed to facilitate the task of manual annotation.
publishDate 2015
dc.date.none.fl_str_mv 2015
2025
2025
dc.type.none.fl_str_mv info:eu-repo/semantics/article
format article
dc.identifier.none.fl_str_mv http://hdl.handle.net/10810/71090
url http://hdl.handle.net/10810/71090
dc.language.none.fl_str_mv Inglés
language_invalid_str_mv Inglés
dc.relation.none.fl_str_mv https://doi.org/10.1093/llc/fqv001
dc.rights.none.fl_str_mv info:eu-repo/semantics/openAccess
© The Author 2015. Published by Oxford University Press on behalf of EADH.
eu_rights_str_mv openAccess
rights_invalid_str_mv © The Author 2015. Published by Oxford University Press on behalf of EADH.
dc.format.none.fl_str_mv application/pdf
dc.publisher.none.fl_str_mv Oxford University Press
publisher.none.fl_str_mv Oxford University Press
dc.source.none.fl_str_mv reponame:Addi. Archivo Digital para la Docencia y la Investigación
instname:Universidad del País Vasco
instname_str Universidad del País Vasco
reponame_str Addi. Archivo Digital para la Docencia y la Investigación
collection Addi. Archivo Digital para la Docencia y la Investigación
repository.name.fl_str_mv
repository.mail.fl_str_mv
_version_ 1869411802685636608
score 15,812455