A Computational Typological Characterization of a Turkish-German Code-Switching
摘要
It is increasingly common to collect linguistic information in unified repositories such as the World Atlas of Language Structures or Grambank. However, these only contain information about languages in the more traditional sense. In other words, other language varieties that do not have a descriptive grammar do not appear. The present study aims to show some of the main syntactic features of the trends detected in Turkish-German Code-Switching: the basic word order, the direction of adpositions, adjectives, adverbs, adverbial clauses, and genitives. To establish a proper comparison with German and Turkish data, 14 different corpora corresponding to all corpora available for these three language varieties in Universal Dependencies 2.11 are studied. To extract the data from the Universal Dependencies corpora, we use the Grew-Match tool. Once the quantitative occurrences have been obtained, we process them and show the quantitative results that allow us to know in a more fine-grained way the behavior of these structures in German, Turkish, and Turkish-German code-switching. We also propose a linguistic categorization through quantificational conversion following. The study points out that the behavior of this Code-Switching is not oriented toward either of its two source languages. The basic Word order is more similar to Turkish (SOV), as is the functioning of genitives. In other aspects such as the order of adposition or modifying adverbs or adverbial clauses, it is more similar to German.