QR-Code

Are the existing training corpora unnecessarily large?

This paper addresses the problem of optimizing the training treebank data because the size and quality of the data has always been a bottleneck for the purposes of training. In previous studies we realized that current corpora used for training machine learning{based dependency parsers contain a sig...

Ausführliche Beschreibung

Gespeichert in:
Bibliografische Detailangaben
Veröffentlicht in:Procesamiento del Lenguaje Natural
Hauptverfasser: Miguel Ballesteros, Jesús Herrera, Virginia Francisco, Pablo Gervás
Format: Artigo
Sprache:Inglês
Veröffentlicht: Sociedad Española para el Procesamiento del Lenguaje Natural 2012
Schlagworte:
Online-Zugang:https://www.redalyc.org/articulo.oa?id=515751748002
Tags: Tag hinzufügen
Keine Tags, Fügen Sie das erste Tag hinzu!