<?xml version='1.0' encoding='UTF-8'?><codeBook xmlns="ddi:codebook:2_5" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="ddi:codebook:2_5 https://ddialliance.org/Specification/DDI-Codebook/2.5/XMLSchema/codebook.xsd" version="2.5"><docDscr><citation><titlStmt><titl>Replication Data for: Análisis contrastivo de los marcadores pragmáticos de vaguedad es que y en plan en el español coloquial actual: indexicalidad social y microhistoria</titl><IDNo agency="DOI">doi:10.18710/I8KK9V</IDNo></titlStmt><distStmt><distrbtr source="archive">DataverseNO</distrbtr><distDate>2026-02-23</distDate></distStmt><verStmt source="archive"><version date="2026-02-23" type="RELEASED">1</version></verStmt><biblCit>Van Den Driessche, Nele; Enghels, Renata, 2026, "Replication Data for: Análisis contrastivo de los marcadores pragmáticos de vaguedad es que y en plan en el español coloquial actual: indexicalidad social y microhistoria", https://doi.org/10.18710/I8KK9V, DataverseNO, V1</biblCit></citation></docDscr><stdyDscr><citation><titlStmt><titl>Replication Data for: Análisis contrastivo de los marcadores pragmáticos de vaguedad es que y en plan en el español coloquial actual: indexicalidad social y microhistoria</titl><altTitl>Replication Data for: Contrastive analysis of the pragmatic markers of vagueness «es que» ('it is that') and «en plan» ('like') in present-day colloquial Spanish: social indexicality and microhistory</altTitl><IDNo agency="DOI">doi:10.18710/I8KK9V</IDNo></titlStmt><rspStmt><AuthEnty affiliation="Ghent University">Van Den Driessche, Nele</AuthEnty><AuthEnty affiliation="Ghent University">Enghels, Renata</AuthEnty><othId role="Data Collector">Nele Van Den Driessche</othId></rspStmt><prodStmt><producer>Ghent University</producer><prodDate>2025-12-18</prodDate><software version="2511">MS Excel</software><software version="2511">Vassarstats</software></prodStmt><distStmt><distrbtr source="archive">DataverseNO</distrbtr><distrbtr abbr="TROLLing" URI="https://trolling.uit.no/">The Tromsø Repository of Language and Linguistics (TROLLing)</distrbtr><contact affiliation="Ghent University" email="Nele.VanDenDriessche@UGent.be">Van Den Driessche, Nele</contact><depositr>Van Den Driessche, Nele</depositr><depDate>2025-12-18</depDate></distStmt><holdings URI="https://doi.org/10.18710/I8KK9V"/></citation><stdyInfo><subject><keyword xml:lang="en">Arts and Humanities</keyword><keyword>social indexicality</keyword><keyword>pragmatic markers</keyword><keyword>colloquial Spanish</keyword><keyword>microdiachrony</keyword><keyword>youth language</keyword></subject><abstract date="2025-12-18">&lt;strong>Dataset abstract&lt;/strong>
&lt;br>This data forms the basis for the research presented in the article "Análisis contrastivo de los marcadores pragmáticos de vaguedad &lt;i>es que&lt;/i> y &lt;i>en plan&lt;/i> en el español coloquial actual: indexicalidad social y microhistoria". It consists of two datafiles. 
&lt;p>The first datafile includes 641 cases of &lt;i>en plan&lt;/i> and 4366 cases of &lt;i>es que&lt;/i>, extracted from CORMA, a conversational corpus of peninsular Spanish compiled between 2016 and 2019. The data were subsequently analysed to investigate the social indexicality of the pragmatic markers.
The second datafile includes 437 cases of &lt;i>en plan&lt;/i> and 1622 cases of &lt;i>es que&lt;/i> from a sample of conversations from CORMA, as well as 137 cases of &lt;i>en plan&lt;/i> and 854 cases of &lt;i>es que&lt;/i> from a sample of conversations from the COLAm corpus. COLAm is a conversational corpus of peninsular Spanish compiled between 2002 and 2007. It consists exclusively of conversations between teenagers. &lt;/p>&lt;p>These data were collected in order to examine the pragmatic function of the pragmatic markers as well as changes in the use and diffusion of the two pragmatic markers in the 21st century. &lt;/p>&lt;/br>

&lt;br>&lt;strong>Abstract of related publication&lt;/strong>&lt;/br>
This article explores two markers of vagueness: &lt;i>es que&lt;/i> (‘it is that’) and &lt;i>en plan&lt;/i> (‘like’), and the role of social indexicality in their use and spread in the 21st century. Through the analysis of colloquial conversations from the CORMA corpus, it is revealed that the productivity of these markers is closely associated with sociolinguistic factors, particularly age. Furthermore, the findings indicate that &lt;i>es que&lt;/i> and &lt;i>en plan&lt;/i> operate at different levels of indexicality: while &lt;i>es que&lt;/i> belongs to the first order, &lt;i>en plan&lt;/i> belongs to the third order, due to its emblematic status in youth language. These levels also account for differences in the dissemination of the markers. A microdiachronic comparison between early 21st century youth language (COLAm) and contemporary youth speech (the youth sub-corpus of CORMA) reveals that third-order elements such as &lt;i>en plan&lt;/i> undergo faster changes and wider spread than first-order elements such as &lt;i>es que&lt;/i>. Moreover, depending on the level, the links between social groups and the frequency of marker use do not always remain stable over time. Additionally, socio-cultural changes are also reflected in linguistic changes. Thus, more traditional sociolinguistic parameters, such as gender and the impact of the network of the school, become less decisive in current youth language under the pressures of globalization and the expansion of social networks. Finally, the microdiachronic study indicates that the pragmatic-functional profile of both markers have remained relatively constant throughout the 21st century.</abstract><sumDscr><timePrd cycle="P1" event="start" date="2002-01-01">2002-01-01</timePrd><timePrd cycle="P1" event="end" date="2007-12-31">2007-12-31</timePrd><timePrd cycle="P2" event="start" date="2016-01-01">2016-01-01</timePrd><timePrd cycle="P2" event="end" date="2019-12-31">2019-12-31</timePrd><collDate cycle="P1" event="start" date="2020-01-01">2020-01-01</collDate><collDate cycle="P1" event="end" date="2024-09-30">2024-09-30</collDate><nation>Spain</nation><geogCover>Madrid</geogCover><dataKind>Corpus data</dataKind></sumDscr></stdyInfo><method><dataColl><collMode>Manual selection of corpus data</collMode><sources><dataSrc>&lt;p>The data contained in this dataset originate from the following sources:&lt;/p>
&lt;ul>
&lt;li>COLAm: Corpus Oral de Lenguaje Adolescente - Madrid. &lt;a href="http://hdl.handle.net/11495/D98E-D689-6A14-5">http://hdl.handle.net/11495/D98E-D689-6A14-5&lt;/a>. COLAm was reused under a CLARIN ACA-NC license (&lt;a href="https://www.kielipankki.fi/support/clarin-eula/">https://www.kielipankki.fi/support/clarin-eula/&lt;/a>). &lt;/li>
&lt;li>CORMA: Corpus Oral de Madrid, Enghels R., De Latte F., L. Roels, N. Van Den Driessche. &lt;a href="https://doi.org/10.5281/zenodo.17455997">https://doi.org/10.5281/zenodo.17455997&lt;/a>.
&lt;/li>
&lt;/ul>

&lt;p>The extracted fragments that are contained in the data files of this dataset only contain annotations based on searches in these sources, but no actual text extracts. They represent non-substantial portions of the sources listed above. Therefore, the reuse (including redistribution) of these data is permitted by the exceptions rules in IPR and database protection regulations, such as Fair use (cf. &lt;a href="https://www.copyright.gov/fair-use/more-info.html">US Copyright Act&lt;/a>), the &lt;a href="http://data.europa.eu/eli/dir/1996/9/oj">EU Database Directive&lt;/a> (cf. art 8 Rights and obligations of lawful users), and the Norwegian Copyright Act (cf. &lt;a href="https://lovdata.no/lov/2018-06-15-40/§24">§ 24 Eneretten til databaser&lt;/a>).&lt;/p></dataSrc></sources><cleanOps>Consistency checking</cleanOps></dataColl><anlyInfo/></method><dataAccs><setAvail/><useStmt/><notes type="DVN:TOU" level="dv">&lt;a href="http://creativecommons.org/publicdomain/zero/1.0">CC0 1.0&lt;/a></notes></dataAccs><othrStdyMat><relPubl><citation><titlStmt><titl>Van Den Driessche, N. &amp; Enghels, R. (2025). Análisis contrastivo de los marcadores pragmáticos de vaguedad 'es que' y 'en plan' en el español coloquial actual: indexicalidad social y microhistoria. Revista Signos. Estudios De Lingüística, 58(119): 548-581.</titl><IDNo agency="doi">10.4151/S0718-09342025011901345</IDNo></titlStmt><biblCit>Van Den Driessche, N. &amp; Enghels, R. (2025). Análisis contrastivo de los marcadores pragmáticos de vaguedad 'es que' y 'en plan' en el español coloquial actual: indexicalidad social y microhistoria. Revista Signos. Estudios De Lingüística, 58(119): 548-581.</biblCit></citation><ExtLink URI="https://doi.org/10.4151/S0718-09342025011901345"/></relPubl></othrStdyMat></stdyDscr><otherMat ID="f268774" URI="https://dataverse.no/api/access/datafile/268774" level="datafile"><labl>00_ReadMe_Esque_Enplan.txt</labl><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/plain</notes></otherMat><otherMat ID="f267856" URI="https://dataverse.no/api/access/datafile/267856" level="datafile"><labl>01_Datafile1_Esque_Enplan_Indexicality.csv</labl><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/comma-separated-values</notes></otherMat><otherMat ID="f267854" URI="https://dataverse.no/api/access/datafile/267854" level="datafile"><labl>02_Datafile2_Esque_Enplan_Microdiachrony_Pragmatics.csv</labl><notes level="file" type="DATAVERSE:CONTENTTYPE" subject="Content/MIME Type">text/comma-separated-values</notes></otherMat></codeBook>