This corpus-based study examines whether written Korean exhibits a syllable-internal sequential asymmetry and, if so, whether the pattern is analogous to the probabilistic C1V bias reported for spoken Korean(Park, 2023). Using a dataset of mono- and disyllabic nouns extracted from the Written Corpus, this study measured contingency between adjacent graphemes within C1VC2 syllables with phi coefficients and analyzed the distributions of the C1V and VC2 sequences. The results provide no distributional evidence for either a C1V bias or a VC2 bias, suggesting that the C1V bias observed in spoken Korean may not be modality-neutral but may be sensitive to the distributional ecologies of written and spoken input. This study extends the usage-based perspective to the written modality by treating written language as a potential input domain for distributional learning and contributes to a more refined account of structural emergence.
목차
Abstract 1. Introduction 2. Distributional Learning and Linguistic Modality 2.1. Distributional Information and Statistical Learning across Modalities 2.2. Cross-modal Differences in Input Distribution and Processing 2.3. Quantification of Contingency: Phi Coefficient and Type Frequency 3. Method 3.1. Dataset Construction and Preprocessing 3.2. Measures of Cohesion 4. Results 5. Discussion 6. Conclusion References