Comparison and Linguistic of Cross-Language English Corpus Driven by Big Data
摘要
The demands of linguistic study and application cannot be satisfied by traditional cross-language English corpora in terms of data acquisition accuracy, diversity, and efficiency. In order to examine the differences in language application, this article contrasts traditional cross-language English corpora with big data-driven cross-language English corpora. By examining the effectiveness of data collecting and translation accuracy, this article thoroughly compares and contrasts traditional corpora with big data-driven corpora from several angles. Comparing five English texts from traditional corpora to five English texts from big data-driven corpora, the average accuracy of the former is 12% lower. The findings of this study offer fresh perspectives and avenues for further language study and application. They also have favorable implications for advancing intercultural and interlingual communication as well as the advancement of machine translation and other technologies.