Image First or Text First? Optimising the Sequencing of Modalities in Large Language Model Prompting and Reasoning Tasks
Our study investigates how the sequencing of text and image inputs within multi-modal prompts affects the reasoning performance of Large Language Models (LLMs). Through empirical evaluations of three major commercial LLM vendors—OpenAI, Google, and Anthropic—alongside a user study on interaction str...
Guardat en:
| Autors principals: | , |
|---|---|
| Format: | Artigo |
| Idioma: | Inglês |
| Publicat: |
MDPI AG
2025-06-01
|
| Col·lecció: | Big Data and Cognitive Computing |
| Matèries: | |
| Accés en línia: | https://www.mdpi.com/2504-2289/9/6/149 |
| Etiquetes: |
Sense etiquetes, Sigues el primer a etiquetar aquest registre!
|
