QR code

LLM-as-a-Judge: automated evaluation of search query parsing using large language models

IntroductionThe adoption of Large Language Models (LLMs) in search systems necessitates new evaluation methodologies beyond traditional rule-based or manual approaches.MethodsWe propose a general framework for evaluating structured outputs using LLMs, focusing on search query parsing within an onlin...

Volledige beschrijving

Bewaard in:
Bibliografische gegevens
Hoofdauteurs: Mehmet Selman Baysan, Serkan Uysal, İrem İşlek, Çağla Çığ Karaman, Tunga Güngör
Formaat: Artigo
Taal:Inglês
Gepubliceerd in: Frontiers Media S.A. 2025-07-01
Reeks:Frontiers in Big Data
Onderwerpen:
Online toegang:https://www.frontiersin.org/articles/10.3389/fdata.2025.1611389/full
Tags: Voeg label toe
Geen labels, Wees de eerste die dit record labelt!