Text readability is vital for effective communication and learning, especially for those with lower information literacy. This research aims to assess Llama 3’s ability to grade readability and compare its alignment with established metrics. For that purpose, we create a new dataset of article lead sections from English and Simple English Wikipedia, covering nine categories. The model is prompted to rate the readability of the texts on a grade-level scale, and an in-depth analysis of the results is conducted. While Llama 3 correlates strongly with most metrics, it may underestimate text grade levels.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Can Llama 3 Accurately Assess Readability? A Comparative Study Using Lead Sections from Wikipedia

  • José Frederico Rodrigues,
  • Henrique Lopes Cardoso,
  • Carla Teixeira Lopes

摘要

Text readability is vital for effective communication and learning, especially for those with lower information literacy. This research aims to assess Llama 3’s ability to grade readability and compare its alignment with established metrics. For that purpose, we create a new dataset of article lead sections from English and Simple English Wikipedia, covering nine categories. The model is prompted to rate the readability of the texts on a grade-level scale, and an in-depth analysis of the results is conducted. While Llama 3 correlates strongly with most metrics, it may underestimate text grade levels.