Welcome to the new version of CaltechAUTHORS. Login is currently restricted to library staff. If you notice any issues, please email coda@library.caltech.edu
Published August 25, 2021 | Submitted
Report Open

Topological Analysis of Syntactic Structures

Abstract

We use the persistent homology method of topological data analysis and dimensional analysis techniques to study data of syntactic structures of world languages. We analyze relations between syntactic parameters in terms of dimensionality, of hierarchical clustering structures, and of non-trivial loops. We show there are relations that hold across language families and additional relations that are family-specific. We then analyze the trees describing the merging structure of persistent connected components for languages in different language families and we show that they partly correlate to historical phylogenetic trees but with significant differences. We also show the existence of interesting non-trivial persistent first homology groups in various language families. We give examples where explicit generators for the persistent first homology can be identified, some of which appear to correspond to homoplasy phenomena, while others may have an explanation in terms of historical linguistics, corresponding to known cases of syntactic borrowing across different language subfamilies.

Attached Files

Submitted - 1903.05181.pdf

Files

1903.05181.pdf
Files (9.2 MB)
Name Size Download all
md5:8d1e0b9d9802c61d381aec28ccd15c90
9.2 MB Preview Download

Additional details

Created:
August 19, 2023
Modified:
March 5, 2024