<?xml version='1.0' encoding='UTF-8'?><codeBook xmlns="ddi:codebook:2_5" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="ddi:codebook:2_5 https://ddialliance.org/Specification/DDI-Codebook/2.5/XMLSchema/codebook.xsd" version="2.5"><docDscr><citation><titlStmt><titl>Non-Standard Allomorphy in Russian Prefixes: Corpus, Experimental, and Statistical Exploration</titl><IDNo agency="DOI">doi:10.18710/FKJAJR</IDNo></titlStmt><distStmt><distrbtr source="archive">DataverseNO</distrbtr><distDate>2014-08-18</distDate></distStmt><verStmt source="archive"><version date="2023-09-28" type="RELEASED">1</version></verStmt><biblCit>Endresen, Anna, 2014, "Non-Standard Allomorphy in Russian Prefixes: Corpus, Experimental, and Statistical Exploration", https://doi.org/10.18710/FKJAJR, DataverseNO, V1</biblCit></citation></docDscr><stdyDscr><citation><titlStmt><titl>Non-Standard Allomorphy in Russian Prefixes: Corpus, Experimental, and Statistical Exploration</titl><IDNo agency="DOI">doi:10.18710/FKJAJR</IDNo></titlStmt><rspStmt><AuthEnty affiliation="UiT The Arctic University of Norway">Endresen, Anna</AuthEnty></rspStmt><prodStmt><producer abbr="UiT">UiT The Arctic University of Norway</producer><prodDate>2014</prodDate></prodStmt><distStmt><distrbtr source="archive">DataverseNO</distrbtr><distrbtr affiliation="UiT The Arctic University of Norway" URI="https://trolling.uit.no/">The Tromsø Repository of Language and Linguistics (TROLLing)</distrbtr><contact affiliation="UiT The Arctic University of Norway" email="anna.endresen@gmail.com">Endresen, Anna</contact><depDate>2014-08-15</depDate></distStmt><holdings URI="https://doi.org/10.18710/FKJAJR"/></citation><stdyInfo><subject><keyword xml:lang="en">Arts and Humanities</keyword><keyword>prefixes</keyword><keyword>Russian</keyword><keyword>verbs</keyword><keyword>allomorphy</keyword><topcClas vocab="&lt;Field term: Choose one or more>">Field: Morphology</topcClas><topcClas>Field: Semantics</topcClas><topcClas vocab="&lt;Time depth: Choose one or more>">Time-depth: synchronic</topcClas><topcClas vocab="&lt;Topic: Choose one or more>">Topic: affixes</topcClas></subject><abstract date="2014-08-15">Abstract: This dissertation challenges the traditional idealized model of allomorphy by confronting it with comprehensive data on 15 Russian aspectual prefixes (RAZ-, RAS-, RAZO-, S-, SO-, PERE-, PRE-, VZ-, VOZ-, O-, OB-, OBO-, U-, VY-, IZ-) collected from corpus and linguistic experiments. The traditional definition narrows allomorphy down to a mere variation of form where the meaning remains constant and variants are distributed complementarily. My findings show that submorphemic semantic differences and distributional overlap are not uncommon properties of morpheme variants. I suggest that allomorphy is a broader phenomenon that goes beyond the axioms of complementary distribution and identical meaning. I examine non-trivial cases of prefix polysemy and multifactorial conditioning of prefix distribution that make it difficult to assess the traditional criteria for allomorphy. Moreover, I present studies of semantic dissimilation of allomorphs and overlap in distribution that violate the absolute criteria for allomorphic relationship. I take the perspective of Cognitive Linguistics and propose an alternative usage-based model of allomorphy that is flexible enough to capture both standard exemplars and non-standard deviations. This model offers detailed applications of several advanced statistical models that optimize the criteria of both semantic “sameness” and distributional complementarity. According to this model, allomorphy is a scalar relationship between morpheme variants – a relationship that can vary in terms of closeness and regularity. Statistical modeling turns the concept of allomorphy into a measurable and verifiable correspondence of form-meaning variation. This makes it possible to measure semantic simi
larity and divergence and distinguish robust patterns of distribution from random effects.</abstract><abstract date="2014-08-15">The set of files includes tagged databases, their versions used in statistical analyses and R codes for the statistical analyses described in the dissertation.</abstract><sumDscr><geogCover>Russia</geogCover><geogCover>Russia</geogCover><dataKind>corpus</dataKind></sumDscr></stdyInfo><method><dataColl><sources/></dataColl><anlyInfo/></method><dataAccs><setAvail/><useStmt/><notes type="DVN:TOU" level="dv">&lt;a href="http://creativecommons.org/publicdomain/zero/1.0">CC0 1.0&lt;/a></notes></dataAccs><othrStdyMat><relPubl><citation><titlStmt><titl>Endresen, Anna. Non-Standard Allomorphy in Russian Prefixes: Corpus, Experimental, and Statistical Exploration. PhD Thesis. UiT The Arctic University of Norway. https://hdl.handle.net/10037/7098.</titl><IDNo agency="handle">10037/7098</IDNo></titlStmt><biblCit>Endresen, Anna. Non-Standard Allomorphy in Russian Prefixes: Corpus, Experimental, and Statistical Exploration. PhD Thesis. UiT The Arctic University of Norway. https://hdl.handle.net/10037/7098.</biblCit></citation><ExtLink URI="https://hdl.handle.net/10037/7098"/></relPubl></othrStdyMat></stdyDscr><otherMat ID="f154" URI="https://dataverse.no/api/access/datafile/154" level="datafile"><labl>ALL PREFIXES CORPUS FACTITIVE VERBS.xlsx</labl><txt>A corpus-based collection of factitive verbs formed by different prefixes (O-, U-, ZA-, RAZ-, etc.) and from different types of bases (adjectival, nominal, pronominal, adverbial, etc.)</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">application/octet-stream</notes></otherMat><otherMat ID="f152" URI="https://dataverse.no/api/access/datafile/152" level="datafile"><labl>DataOU.csv</labl><txt>This is a reduced version of the database for the purposes of the statistical analysis of the corpus data (155 verbs in O- and U-).</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f142" URI="https://dataverse.no/api/access/datafile/142" level="datafile"><labl>datOB.csv</labl><txt>This is the database of subjects responses for the Linear Regression Mixed Effects Model</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f139" URI="https://dataverse.no/api/access/datafile/139" level="datafile"><labl>ObCorpus.csv</labl><txt>This is the database for the statistical analysis of the corpus data on O-, OB-, and OBO-. It excludes deetymologized verbs.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f141" URI="https://dataverse.no/api/access/datafile/141" level="datafile"><labl>ObExperimentRandomForestData.csv</labl><txt>This is the database of subjects responses for the Classification Tree &amp; Random Forests model</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f144" URI="https://dataverse.no/api/access/datafile/144" level="datafile"><labl>ObExperimentSubjects.csv</labl><txt>This spreadsheet contains anonymous sociolinguistic information about the subjects who participated in the experiment.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f138" URI="https://dataverse.no/api/access/datafile/138" level="datafile"><labl>O OB DATABASE CORPUS.xlsx</labl><txt>This database contains 1037 verbs in O-, OB-, and OBO- collected from the Russian National Corpus.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">application/octet-stream</notes></otherMat><otherMat ID="f151" URI="https://dataverse.no/api/access/datafile/151" level="datafile"><labl>O U CORPUS DATA.xlsx</labl><txt>This database contains 155 factitive (change-of-state) verbs in O- and U- collected from the Russian National Corpus.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">application/octet-stream</notes></otherMat><otherMat ID="f155" URI="https://dataverse.no/api/access/datafile/155" level="datafile"><labl>O U EXPERIMENTAL DATA.xlsx</labl><txt>Acceptability scores elicited from 121 subjects in the experiment on Russian factitive verbs in O- and U-.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">application/octet-stream</notes></otherMat><otherMat ID="f145" URI="https://dataverse.no/api/access/datafile/145" level="datafile"><labl>PERE PRE DATABASE.xlsx</labl><txt>This database contains 945 verbs prefixed in PERE- and PRE- collected from the Russian National Corpus.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">application/octet-stream</notes></otherMat><otherMat ID="f132" URI="https://dataverse.no/api/access/datafile/132" level="datafile"><labl>RAZ DATABASE.xlsx</labl><txt>This database contains 210 perfective Russian verbs formed by the prefixes RAZ-, RAS-, and RAZO- 'apart'. These prefixes represent a case of Standard Allomorphy conditioned by phonological and morphophonological factors. Two phenomena are at work here: voicing assimilation across a prefix-root boundary (prefixes RAZ- ~ RAS- 'apart') and vocalization of consonant-final Russian prefixes (RAZ- ~ RAZO- 'apart').</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">application/octet-stream</notes></otherMat><otherMat ID="f133" URI="https://dataverse.no/api/access/datafile/133" level="datafile"><labl>RAZ_RAS.csv</labl><txt>This is a csv version of the database designed for the purposes of the statistical analysis documented in the R code.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f134" URI="https://dataverse.no/api/access/datafile/134" level="datafile"><labl>RAZ_RAS_RAZO.csv</labl><txt>This is a csv version of the database designed for the purposes of the statistical analysis documented in the R code.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f140" URI="https://dataverse.no/api/access/datafile/140" level="datafile"><labl>R script O OB CORPUS.R</labl><txt>This is the code for the statistical analysis of the corpus data on the prefixes O-, OB-, and OBO-.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f143" URI="https://dataverse.no/api/access/datafile/143" level="datafile"><labl>R script O OB EXPERIMENT.R</labl><txt>This is the R code for both statistical models for the experimental data on the prefixes O-, OB-, and OBO-.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f153" URI="https://dataverse.no/api/access/datafile/153" level="datafile"><labl>R script O U CORPUS DATA.R</labl><txt>R code for the statistical analysis of 155 factitive verbs in O- and U-</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f157" URI="https://dataverse.no/api/access/datafile/157" level="datafile"><labl>R script O U EXPERIMENT ALL MODELS.R</labl><txt>This is the R code for all statistical models applied to the experimental data on O- and U- in factitive verbs.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=UTF-8</notes></otherMat><otherMat ID="f146" URI="https://dataverse.no/api/access/datafile/146" level="datafile"><labl>R script PERE PRE.R</labl><txt>This is the R code for several statistical tests discussed in Chapter 6.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f135" URI="https://dataverse.no/api/access/datafile/135" level="datafile"><labl>R script RAZ.R</labl><txt>This is the R code for the statistical analysis. The statistical analysis models the distribution of these polysemous but standard allomorphs and evaluates the relative impact of each factor.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f137" URI="https://dataverse.no/api/access/datafile/137" level="datafile"><labl>R script S SO.R</labl><txt>This is the R code for the statistical analysis discussed in Chapter 4.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f150" URI="https://dataverse.no/api/access/datafile/150" level="datafile"><labl>R script VY IZ.R</labl><txt>This R code documents the statistical tests discussed in Chapter 8.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f148" URI="https://dataverse.no/api/access/datafile/148" level="datafile"><labl>R script VZ VOZ.R</labl><txt>This R code documents the statistical test discussed in Chapter 7.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f156" URI="https://dataverse.no/api/access/datafile/156" level="datafile"><labl>ScoresForStatistics.csv</labl><txt>This is the database for the statistical analysis.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain; charset=US-ASCII</notes></otherMat><otherMat ID="f136" URI="https://dataverse.no/api/access/datafile/136" level="datafile"><labl>S SO MODERN RUSSIAN.xlsx</labl><txt>This database contains 998 Modern Russian verbs in S- and SO- collected from the Russian National Corpus. Each verb is assigned a number of tags and is illustrated with examples from the corpus.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">application/octet-stream</notes></otherMat><otherMat ID="f149" URI="https://dataverse.no/api/access/datafile/149" level="datafile"><labl>VY IZ DATABASE.xlsx</labl><txt>This database contains 989 verbs in VY- and IZ- attested in the Russian National Corpus.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">application/octet-stream</notes></otherMat><otherMat ID="f147" URI="https://dataverse.no/api/access/datafile/147" level="datafile"><labl>VZ VOZ DATABASE.xls</labl><txt>This database contains 384 verbs in VZ- and VOZ- collected from the Russian National Corpus.</txt><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">application/vnd.ms-excel</notes></otherMat></codeBook>