Supporting materials: Vector-Based Analysis of Phonetic Iconicity in Spatial Dimension Adjectives in Yunnan Minority Languages
This dataset and analysis pipeline support a cross‑linguistic investigation into phonetic iconicity in spatial adjectives across 36 minority languages of Yunnan, China. The data extracted from the Yunnan Provincial Gazetteer: Gazetteer of Ethnic Minority Languages and Scripts (1989). We provide: IPA transcriptions for 18 spatial dimension terms per language. Numerical phonetic vectors derived from articulatory features. Difference vectors for each antonym pair to capture systematic contrasts. Principal Component Analysis (PCA) results. Sonority Sequencing Principle (SSP) data for vowels and consonants. Repository Structure: Data_and_Results/ – Curated results, including visualisations, PCA outputs, and a detailed data catalog. Supporting_Materials/ – Complete computational workflow, comprising: SDAs_Tool/ – All Python scripts, configuration files, and three detailed READMEs. SDAs_Data/ – Auto‑generated intermediate and outputs from the scripts. yunnan_languages_SDAs/ – CLDF‑formatted dataset. soundvectors_Data/ – Segment‑level feature vectors and aggregated statistics. All data are provided under a CC‑BY 4.0 license. Please cite this repository and the original source (Yunnan Provincial Chronicles, 1989) when using the data.
Authors
- Wenqi Li (ORCID: https://orcid.org/0000-0002-7246-0298)
- Yutong Kuang
- Danqing Liu (ORCID: https://orcid.org/0000-0001-8830-0443)
Institutions
- Shenzhen University (CN)
- University of Macau (MO)
Publication Details
- Journal
- Zenodo (CERN European Organization for Nuclear Research)
- Published
- 2026-09-06
- DOI
- https://doi.org/10.5281/zenodo.22541250
- Primary Topic
- Multisensory perception and integration
- Type
- article
- Field-Weighted Citation Impact
- 0.00