Are the existing training corpora unnecessarily large?
This paper addresses the problem of optimizing the training treebank data because the size and quality of the data has always been a bottleneck for the purposes of training. In previous studies we realized that current corpora used for training machine learning{based dependency parsers contain a sig...
Shranjeno v:
| izdano v: | Procesamiento del Lenguaje Natural |
|---|---|
| Principais autores: | , , , |
| Format: | Artigo |
| Jezik: | Inglês |
| Izdano: |
Sociedad Española para el Procesamiento del Lenguaje Natural
2012
|
| Teme: | |
| Online dostop: | https://www.redalyc.org/articulo.oa?id=515751748002 |
| Oznake: |
Brez oznak, prvi označite!
|
