QR kód

Tokenization efficiency of current foundational large language models for the Ukrainian language

Foundational large language models (LLMs) are deployed in multilingual environments across a range of general and narrow task domains. These models generate text token by token, making them slower and more computationally expensive for low-resource languages that are underrepresented in the tokenize...

Celý popis

Uloženo v:
Podrobná bibliografie
Hlavní autoři: Daniil Maksymenko, Oleksii Turuta
Médium: Artigo
Jazyk:Inglês
Vydáno: Frontiers Media S.A. 2025-08-01
Edice:Frontiers in Artificial Intelligence
Témata:
On-line přístup:https://www.frontiersin.org/articles/10.3389/frai.2025.1538165/full
Tagy: Přidat tag
Žádné tagy, Buďte první, kdo vytvoří štítek k tomuto záznamu!