Are the existing training corpora unnecessarily large?
This paper addresses the problem of optimizing the training treebank data because the size and quality of the data has always been a bottleneck for the purposes of training. In previous studies we realized that current corpora used for training machine learning{based dependency parsers contain a sig...
Gespeichert in:
| Veröffentlicht in: | Procesamiento del Lenguaje Natural |
|---|---|
| Hauptverfasser: | , , , |
| Format: | Artigo |
| Sprache: | Inglês |
| Veröffentlicht: |
Sociedad Española para el Procesamiento del Lenguaje Natural
2012
|
| Schlagworte: | |
| Online-Zugang: | https://www.redalyc.org/articulo.oa?id=515751748002 |
| Tags: |
Keine Tags, Fügen Sie das erste Tag hinzu!
|
