← BackMistral AIApr 9, 2025, 20:00 UTC+8Evaluating RAG with LLM as a JudgeMistral 介绍用 LLM as a Judge 评估 RAG 系统,由 judge LLM 按数值、二元或定性量表为 generator LLM 的回答打分,再对评测数据集求加权总分。Read the original ↗ShareCopy linkShare imageExport Markdown