Are the existing training corpora unnecessarily large?
This paper addresses the problem of optimizing the training treebank data because the size and quality of the data has always been a bottleneck for the purposes of training. In previous studies we realized that current corpora used for training machine learning{based dependency parsers contain a sig...
Na minha lista:
| Publicado no: | Procesamiento del Lenguaje Natural |
|---|---|
| Principais autores: | , , , |
| Formato: | Artigo |
| Idioma: | Inglês |
| Publicado em: |
Sociedad Española para el Procesamiento del Lenguaje Natural
2012
|
| Assuntos: | |
| Acesso em linha: | https://www.redalyc.org/articulo.oa?id=515751748002 |
| Tags: |
Sem tags, seja o primeiro a adicionar uma tag!
|
