Size Matters: The Impact of Training Size in Taxonomically-Enriched Word Embeddings
Word embeddings trained on natural corpora (e.g., newspaper collections, Wikipedia or the Web) excel in capturing thematic similarity (“topical relatedness”) on word pairs such as ‘coffee’ and ‘cup’ or ’bus’ and ‘road’. However, they are less successful on pairs showing taxonomic similarity, like ‘c...
Bewaard in:
| Hoofdauteurs: | , , |
|---|---|
| Formaat: | Artigo |
| Taal: | Inglês |
| Gepubliceerd in: |
De Gruyter
2019-10-01
|
| Reeks: | Open Computer Science |
| Onderwerpen: | |
| Online toegang: | https://doi.org/10.1515/comp-2019-0009 |
| Tags: |
Geen labels, Wees de eerste die dit record labelt!
|
