ELOQUENT CLEF Shared Tasks for Evaluation of Generative Language Model Quality, 2025 Edition
摘要
The ELOQUENT lab for evaluation of generative language model quality and usefulness addresses high-level quality criteria through a set of open-ended shared tasks implemented, where possible, to leverage the ability of systems built on language model to assess their own capacity.