The ELOQUENT lab for evaluation of generative language model quality and usefulness addresses high-level quality criteria through a set of open-ended shared tasks implemented, where possible, to leverage the ability of systems built on language model to assess their own capacity.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

ELOQUENT CLEF Shared Tasks for Evaluation of Generative Language Model Quality, 2025 Edition

  • Jussi Karlgren,
  • Ekaterina Artemova,
  • Ondřej Bojar,
  • Vladislav Mikhailov,
  • Magnus Sahlgren,
  • Erik Velldal,
  • Lilja Øvrelid

摘要

The ELOQUENT lab for evaluation of generative language model quality and usefulness addresses high-level quality criteria through a set of open-ended shared tasks implemented, where possible, to leverage the ability of systems built on language model to assess their own capacity.