4,761 to 4,770 of 5,205 Results
Sep 22, 2016 -
Data on Icelandic pre-aspiration
Waveform Audio - 47.1 MB -
MD5: 158e38eed342c9d08f51969f5535b277
|
Sep 1, 2016
Holliday, Jeffrey J.; Turnbull, Rory; Eychenne, Julien, 2016, "K-SPAN (Korean Surface Phones and Neighborhoods)", https://doi.org/10.18710/TWM79F, DataverseNO, V2, UNF:6:NWbRmiBvO5wWcDN2QHCQJw== [fileUNF]
This corpus provides surface phonetic forms derived from a publicly available orthographic corpus of Korean, along with neighborhood density statistics for each word in the corpus. The surface phonetic forms are rendered in an ASCII-encoded scheme, which allows users to explore and query the corpus without having to read Korean orthography. |
Sep 1, 2016 -
K-SPAN (Korean Surface Phones and Neighborhoods)
Tabular Data - 8.6 MB - 19 Variables, 63836 Observations - UNF:6:NWbRmiBvO5wWcDN2QHCQJw==
K-SPAN database |
Sep 1, 2016 -
K-SPAN (Korean Surface Phones and Neighborhoods)
Adobe PDF - 721.0 KB -
MD5: 2bf63cc0704a2fe3b4c40445192af77f
documentation for the K-SPAN database |
Sep 1, 2016 -
K-SPAN (Korean Surface Phones and Neighborhoods)
Python Source Code - 4.5 KB -
MD5: a87dffdad6725b500a187e1168859149
Script to merge K-SPAN with the NIKL corpus (see documentation) |
Jul 2, 2016
Chromý, Jan, 2016, "Data from the project Sociolinguistic analysis of the use of prothetic /v/ in Czech", https://doi.org/10.18710/AGL9FD, DataverseNO, V1
Data from the project Sociolinguistic analysis of the use of prothetic /v/ in Czech. Altogether, 28 893 tokens of words which may contain prothetic v- taken from sociolinguistic interviews with 159 speakers from five Czech cities (Prague, Brno, České Budějovice, Plzeň and Hradec Králové). The speakers are either from younger (20 to 30 years) or old... |
Plain Text - 5.6 MB -
MD5: f1fdf4f8e54a752b9c49f9c57dc8b354
Data set from the project Sociolinguistic analysis of the use of prothetic /v/ in Czech. |
Plain Text - 6.1 KB -
MD5: 79b412139b6008a61b9495d901971300
Description for each variable in the dataset. |
Apr 25, 2016
Pepper, Steve, 2016, "Replication data for: Windmills, Nizaa and the typology of binominal compounds", https://doi.org/10.18710/MP1JF6, DataverseNO, V1
This data set consists of 500+ nominal compounds from the African language Nizaa (sgi; Niger-Congo, Cameroon). It is based on an unpublished word list collected by Rolf Theil ( genannt Endresen) of the University of Oslo in the 1980s. Each compound and its constituents are glossed and annotated for word class, and 201 transparent noun-noun compound... |
Plain Text - 66.3 KB -
MD5: 9e421ee0c4c69f7b89fdbeca76fa22b8
Data set in UTF-8, tab-delimited format |
