Cargando…

Scaling behavior and text cohesion in Korean texts

This study examines whether different types of texts, particularly in Korean, can be distinguished by the scaling exponent and degree of text cohesion. We use the controlled growth process model to incorporate the interaction effect into a power-law distribution and estimate the implied parameter ex...

Descripción completa

Detalles Bibliográficos
Autores principales: Kim, Hokyun, Park, Sanghu, Jeong, Minhyuk, Byun, Hyungi, Kim, Juyub, Lee, Doo Yong, Jeon, Jooyoung, Yi, Eojin, Ahn, Kwangwon
Formato: Online Artículo Texto
Lenguaje:English
Publicado: Public Library of Science 2023
Materias:
Acceso en línea:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC10470962/
https://www.ncbi.nlm.nih.gov/pubmed/37651361
http://dx.doi.org/10.1371/journal.pone.0290168
Descripción
Sumario:This study examines whether different types of texts, particularly in Korean, can be distinguished by the scaling exponent and degree of text cohesion. We use the controlled growth process model to incorporate the interaction effect into a power-law distribution and estimate the implied parameter explaining the degree of text cohesiveness in a word distribution. We find that the word distributions of Korean languages differ from English regarding the range of scaling exponents. Additionally, different types of Korean texts display similar scaling exponents regardless of their genre. However, the interaction effect is higher for expert reports than for the benchmark novels. The findings suggest a valid framework for explaining the scaling phenomena of word distribution based on microscale interactions. It also suggests that a viable method exists for inferring text genres based on text cohesion.