<?xml version='1.0' encoding='UTF-8'?><codeBook xmlns="ddi:codebook:2_5" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="ddi:codebook:2_5 https://ddialliance.org/Specification/DDI-Codebook/2.5/XMLSchema/codebook.xsd" version="2.5"><docDscr><citation><titlStmt><titl>Supporting Data for: Investigating the covariation of features in ‘crowded’ vowel spaces – A corpus-based study of BOOT and BIT in Standard Scottish English</titl><IDNo agency="DOI">doi:10.18710/EYUPME</IDNo></titlStmt><distStmt><distrbtr source="archive">DataverseNO</distrbtr><distDate>2026-08-26</distDate></distStmt><verStmt source="archive"><version date="2026-08-26" type="RELEASED">1</version></verStmt><biblCit>Schützler, Ole, 2026, "Supporting Data for: Investigating the covariation of features in ‘crowded’ vowel spaces – A corpus-based study of BOOT and BIT in Standard Scottish English", https://doi.org/10.18710/EYUPME, DataverseNO, V1</biblCit></citation></docDscr><stdyDscr><citation><titlStmt><titl>Supporting Data for: Investigating the covariation of features in ‘crowded’ vowel spaces – A corpus-based study of BOOT and BIT in Standard Scottish English</titl><IDNo agency="DOI">doi:10.18710/EYUPME</IDNo></titlStmt><rspStmt><AuthEnty affiliation="Leipzig University">Schützler, Ole</AuthEnty><othId role="Data Collector">Gut, Ulrike</othId><othId role="Data Curator">Li, Zeyu</othId></rspStmt><prodStmt><producer>Leipzig University</producer><prodPlac>Scotland, UK</prodPlac><software version="6.1.51">Praat</software><software version="2408, Build 16.0.17932.20286">Microsoft Excel</software><software version="4.1.1">R</software><software version="2025.05.1">RStudio</software><software version="2.16.1">R package 'brms'</software><software version="0.20-44">R package 'lattice'</software><software version="0.6-29">R package 'latticeExtra'</software><grantNo agency="German Research Foundation (DFG)">SCHU 3250/1-1</grantNo></prodStmt><distStmt><distrbtr source="archive">DataverseNO</distrbtr><distrbtr abbr="TROLLing" URI="https://trolling.uit.no/">The Tromsø Repository of Language and Linguistics (TROLLing)</distrbtr><contact affiliation="Leipzig University" email="ole.schuetzler@uni-leipzig.de">Schützler, Ole</contact><depositr>Schützler, Ole</depositr><depDate>2025-07-02</depDate></distStmt><holdings URI="https://doi.org/10.18710/EYUPME"/></citation><stdyInfo><subject><keyword xml:lang="en">Arts and Humanities</keyword><keyword>Standard Scottish English</keyword><keyword>acoustic phonetics</keyword><keyword>sociophonetics</keyword><keyword>Dispersion Theory</keyword><keyword>vowel variation</keyword><keyword>corpus data</keyword></subject><abstract date="2026-08-26">The study focuses on the vowel categories BOOT and BIT in Standard Scottish English (SSE). The BOOT-vowel tends to be relatively front, the BIT-vowel tends to be lowered and/or centralised, and both are subject to sociolinguistic variation. In addition, this paper argues that they are also competing for limited articulatory (and acoustic) space, so that socially motivated variation may be constrained by the need for vowel categories to remain distinct. The acoustic study that is presented explores the sociolinguistic variation of the two vowels against the background of systemic, space-related constraints. The centre frequencies of the first two formants of all monophthong vowels are measured in &lt;i>n&lt;/i> = 46 recordings from the Scottish component of the International Corpus of English. Measurements are normalised and the variation of BOOT and BIT is modelled statistically, using linear mixed-effects regression; additionally, the correlation of random intercepts for speakers are inspected. There are virtually no age and gender effects on the variation of the BOOT-vowel, while younger age and male gender correlate with more centralised variants of the BIT-vowel. These factors apart, there is evidence of a tendency for the two vowels to covary in a way that maintains their distance from each other in acoustic space.</abstract><sumDscr><timePrd cycle="P1" event="start" date="2012-01-01">2012-01-01</timePrd><timePrd cycle="P1" event="end" date="2020-12-31">2020-12-31</timePrd><collDate cycle="P1" event="start" date="2024-07-01">2024-07-01</collDate><collDate cycle="P1" event="end" date="2024-09-01">2024-09-01</collDate><nation>United Kingdom</nation><geogCover>Scotland</geogCover><dataKind>Annotated acoustic measurements</dataKind></sumDscr></stdyInfo><method><dataColl><collectorTraining>trained linguist with a focus on phonology/phonetics</collectorTraining><collMode>A phonological transcription automatically generated via forced alignment and manually corrected by the corpus compilers was customised by the data collector (e.g. ensuring that the focus was on stressed contexts, checking for remaining errors). Measuring intervals for relevant vowel categories were set manually. The extraction of formant readings was then done automatically.</collMode><sources><dataSrc>&lt;p>The data in this dataset were retrieved from the 
&lt;a href="https://github.com/langres/ICE-Scotland" target="_blank">International Corpus of English: Scottish Component&lt;/a>, 
under the Creative Commons CC BY-NC-SA 3.0 DE license, and from the British National Corpus available at
&lt;a href="https://www.english-corpora.org/bnc/" target="_blank">https://www.english-corpora.org/bnc/&lt;/a>, licensed for academic and non-commercial purposes.&lt;/p>
&lt;p>The extracted text fragments that are contained in the data files SSE_vowels.csv,  SSE_I.csv,  SSE_U.csv of this dataset only represent non-substantial portions of the sources listed above, and they do not represent coherent larger texts. Therefore, the reuse (including redistribution) of these excerpts is permitted by the exceptions rules in IPR and database protection regulations, such as Fair use (USA cf. &lt;a href="https://www.copyright.gov/fair-use/more-info.html" title="Fair use" target="_blank">US Copyright Act&lt;/a>), Fair dealing (UK; cf. &lt;a href="https://www.gov.uk/guidance/exceptions-to-copyright" title="Fair dealing" target="_blank">Exceptions to copyright&lt;/a>), the &lt;a href="http://data.europa.eu/eli/dir/1996/9/2019-06-06" title="Lawful users" target="_blank">EU Database Directive&lt;/a> (cf. article 8 Rights and obligations of lawful users), "lover, forskrifter, rettsavgjørelser og andre vedtak av offentlig myndighet" (Norway; cf. &lt;a href="https://lovdata.no/lov/2018-06-15-40/§14" title="offentlige vedtak" target="_blank">§ 14 in Åndsverkloven&lt;/a>), "uvesentlige deler av databaser" (Norway; cf. &lt;a href="https://lovdata.no/lov/2018-06-15-40/§24" title="uvesentlige deler av databaser" target="_blank">§ 24 in Åndsverkloven&lt;/a>), "sitatretten" (Norway; cf. &lt;a href="https://lovdata.no/lov/2018-06-15-40/§29" title="sitatretter" target="_blank">§ 29 in Åndsverkloven&lt;/a>). As these excerpts do not represent substantial parts of the reused sources, the redistribution of these excerpts is according to Creative Commons (CC) also permitted if they are extracted from sources that are distributed under Creative Commons licenses (cf. question "Do I always have to comply with the license terms? If not, what are the exceptions?" in the &lt;a href="https://creativecommons.org/faq/" title="CC FAQ" target="_blank">Creative Commons Frequently Asked Questions&lt;/a>). Additionally, the use and redistribution of the frequency information obtained from the British National Corpus available at &lt;a href="https://www.english-corpora.org/bnc/" target="_blank">https://www.english-corpora.org/bnc/&lt;/a>, is permitted by the exceptions rules in IPR and database protection regulations mentioned above.&lt;/p></dataSrc><srcDocu>&lt;ul>
  &lt;li>
    &lt;a href="https://www.ice-corpora.uzh.ch/en.html" target="_blank">ICE-website at the University of Zurich&lt;/a>
  &lt;/li>
  &lt;li>
    Davies, Mark. 2004. &lt;i>British National Corpus&lt;/i> (from Oxford University Press). Available online at 
    &lt;a href="https://www.english-corpora.org/bnc/" target="_blank">https://www.english-corpora.org/bnc/&lt;/a>.
  &lt;/li>
  &lt;li>
    Kirk, John &amp; Gerald Nelson. 2018. The International Corpus of English project: A progress report. 
    &lt;i>World Englishes&lt;/i> 34(4): 697–716. 
    &lt;a href="https://doi.org/10.1111/weng.12350" target="_blank">https://doi.org/10.1111/weng.12350&lt;/a>
  &lt;/li>
  &lt;li>
    Nelson, Gerald, Sean Wallis &amp; Bas Aarts. 2002. 
    &lt;i>Exploring Natural Language. Working with the British Component of the International Corpus of English.&lt;/i> 
    Amsterdam: John Benjamins.
  &lt;/li>
  &lt;li>
    Schützler, Ole, Ulrike Gut &amp; Robert Fuchs. 2017. 
    New perspectives on Scottish Standard English: Introducing the Scottish component of the International Corpus of English. 
    In Joan C. Beal &amp; Sylvie Hancil (Hg.), &lt;i>Perspectives on northern Englishes.&lt;/i> Berlin: Mouton de Gruyter [TiEL 96]. 273–302. 
    &lt;a href="https://doi.org/10.1515/9783110450903-012" target="_blank">https://doi.org/10.1515/9783110450903-012&lt;/a>
  &lt;/li>
  &lt;li>
    Access to corpus data: 
    &lt;a href="https://github.com/langres/ICE-Scotland" target="_blank">https://github.com/langres/ICE-Scotland&lt;/a>
  &lt;/li>
&lt;/ul></srcDocu></sources><collSitu>The original sampling of texts for ICE-Scotland followed the general sampling scheme for corpora from that family and depended on the accessibility of institutions and the cooperation of individuals. It can therefore not be considered random.</collSitu><actMin>Measuring intervals were set manually, not automatically, taking into account spectrographic information. This way, many 'difficult' tokens could be salvaged.</actMin><cleanOps>Tokens classified as unmeasurable (due to unclear or extremely unstable formant patterns or spurious formants) were excluded when setting the measuring intervals.</cleanOps></dataColl><anlyInfo/></method><dataAccs><setAvail/><useStmt/><notes type="DVN:TOU" level="dv">&lt;p>With the exception of the tabular files  SSE_vowels.csv, SSE_I.csv, SSE_U.csv, the dataset " Investigating the covariation of features in ‘crowded’ vowel spaces – A corpus-based study of /u/ and /ɪ/ in Standard Scottish English: Acoustic measurements " has been marked as dedicated to the public domain, as described here: &lt;a href="https://creativecommons.org/publicdomain/zero/1.0/">https://creativecommons.org/publicdomain/zero/1.0/&lt;/a>.&lt;/p>  

&lt;p>The tabular files SSE_vowels.csv, SSE_I.csv, SSE_U.csv contain source information (i.e. the columns “text” and “textlength”), speaker information (i.e. the columns “gender” and “age”), and text fragments (i.e. the columns “lemma” and “word”) that have been extracted from the International Corpus of English: Scottish Component, available at &lt;a href="https://github.com/langres/ICE-Scotland">github.com/langres/ICE-Scotland&lt;/a>, licensed under CC BY-NC-SA 3.0 DE, under limitations and exceptions to IPR and database protection regulations.&lt;/p> 

&lt;p>Additionally, these tabular files contain analyses (i.e. the columns “spoBNC_raw”, “spoBNC_pmw” and “spoBNC_log_pmw”) that have been based on the British National Corpus (BNC), available at &lt;a href="https://www.english-corpora.org/bnc/">english-corpora.org/bnc/&lt;/a>, licensed for academic and non-commercial purposes, is permitted under the &lt;a href="https://fairuse.stanford.edu/overview/fair-use/" target="_blank">US Fair Use Law&lt;/a>.&lt;/p> 

&lt;p> The contribution of the author of the present dataset to these files, as detailed in the ReadMe file, is licensed under the Creative Commons Attribution 4.0 International (CC BY 4.0) license, as described here: &lt;a href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/&lt;/a>.&lt;/p> </notes></dataAccs><othrStdyMat><relMat>Documentation of analysis: &lt;a href="https://osf.io/fs3nk/" target="_blank">https://osf.io/fs3nk/&lt;/a></relMat><relPubl><citation><titlStmt><titl>Schützler, Ole. 2026. "Investigating the covariation of features in ‘crowded’ vowel spaces: A corpus-based study of BOOT and BIT in Standard Scottish English." In Philipp Meer &amp;  Ulrike Gut (eds.), &lt;i>English Corpus Phonetics and Phonology: Current Approaches and Future Directions&lt;/i> (Topics in English Linguistics 122), 95-127. Berlin: de Gruyter Mouton.</titl><IDNo agency="doi">10.1515/9783112213254-004</IDNo></titlStmt><biblCit>Schützler, Ole. 2026. "Investigating the covariation of features in ‘crowded’ vowel spaces: A corpus-based study of BOOT and BIT in Standard Scottish English." In Philipp Meer &amp;  Ulrike Gut (eds.), &lt;i>English Corpus Phonetics and Phonology: Current Approaches and Future Directions&lt;/i> (Topics in English Linguistics 122), 95-127. Berlin: de Gruyter Mouton.</biblCit></citation><ExtLink URI="https://doi.org/10.1515/9783112213254-004"/></relPubl></othrStdyMat></stdyDscr><otherMat ID="f295900" URI="https://dataverse.no/api/access/datafile/295900" level="datafile"><labl>00_ReadMe_CovariationOfFeaturesInCrowdedVowelSpaces.txt</labl><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain</notes></otherMat><otherMat ID="f254617" URI="https://dataverse.no/api/access/datafile/254617" level="datafile"><labl>Licence_ICE-SCO.pdf</labl><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">application/pdf</notes></otherMat><otherMat ID="f254615" URI="https://dataverse.no/api/access/datafile/254615" level="datafile"><labl>SSE_I.csv</labl><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/comma-separated-values</notes></otherMat><otherMat ID="f254700" URI="https://dataverse.no/api/access/datafile/254700" level="datafile"><labl>SSE_U.csv</labl><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/comma-separated-values</notes></otherMat><otherMat ID="f254618" URI="https://dataverse.no/api/access/datafile/254618" level="datafile"><labl>SSE_vowels.csv</labl><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/comma-separated-values</notes></otherMat></codeBook>