The development of Large Language models, such as GPT, has already revolutionized the fields of Artificial Intelligence and Natural Language Processing, enabling human-like conversation generation. The increasing diffusion of ChatGPT is expected to influence many more areas and society as a whole. Education is one of the domains which is probably going to be affected the most by this revolution. Therefore, analyzing the general public’s opinions on emerging technologies, such as ChatGPT, becomes crucial to better understand their potential ethical implications and applications within the educational landscape, and might be a precious resource for various purposes, including performance evaluation, and identification of needs and expectations. This paper presents a novel dataset comprising 236 thousand tweets capturing public discourse surrounding ChatGPT and Education, along with enriched dimensions extracted from the original texts. By leveraging data enrichment techniques, we make the dataset accessible for analysis by researchers and practitioners who may not have expertise in programming and Natural Language Processing. This dataset serves as a valuable resource for performing exploratory data analysis about Twitter users’ perceptions of the potential impact of ChatGPT on the educational processes.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

The ChatGPT and Education Tweets Dataset

  • Simone Barandoni,
  • Filippo Chiarello,
  • Vito Giordano,
  • Gualtiero Fantoni

摘要

The development of Large Language models, such as GPT, has already revolutionized the fields of Artificial Intelligence and Natural Language Processing, enabling human-like conversation generation. The increasing diffusion of ChatGPT is expected to influence many more areas and society as a whole. Education is one of the domains which is probably going to be affected the most by this revolution. Therefore, analyzing the general public’s opinions on emerging technologies, such as ChatGPT, becomes crucial to better understand their potential ethical implications and applications within the educational landscape, and might be a precious resource for various purposes, including performance evaluation, and identification of needs and expectations. This paper presents a novel dataset comprising 236 thousand tweets capturing public discourse surrounding ChatGPT and Education, along with enriched dimensions extracted from the original texts. By leveraging data enrichment techniques, we make the dataset accessible for analysis by researchers and practitioners who may not have expertise in programming and Natural Language Processing. This dataset serves as a valuable resource for performing exploratory data analysis about Twitter users’ perceptions of the potential impact of ChatGPT on the educational processes.