None
TL
The Indo-European Cognate Relationships Dataset.
['David Eddyshaw']
languagehat.com
Matthew Scarborough has featured at LH many times (see, e.g., here), and he has now posted The Indo-European Cognate Relationships dataset (Scientific Data 12.
The paper discusses the underlying dataset, and its organisation and structure and is published together with a revised version (v.1.2) of the dataset on Zenodo.
From the article’s abstract:The Indo-European Cognate Relationships (IE-CoR) dataset is an open-access relational dataset showing how related, inherited words (‘cognates’) pattern across 160 languages of the Indo-European family.
IE-CoR is intended as a benchmark dataset for computational research into the evolution of the Indo-European languages.
The dataset is extendable to further languages and meanings and follows the Cross-Linguistic Data Format (CLDF) protocols for linguistic data.
['used'
'language'
'dataset'
'indoeuropean'
'data'
'relationships'
'languages'
'finally'
'cognate'
'languagehatcom'
'paper']